GitHub/Anki - OBJNULLs Forgejo: Beyond coding. We Forge.

mirror of https://github.com/ankitects/anki.git synced 2025-09-23 08:22:24 -04:00

Author	SHA1	Message	Date
Damien Elmes	83f8ef45ff	anki2 importing and reorganize import code	2011-10-21 07:36:44 +09:00
Damien Elmes	362ae3eee2	initial work on sync refactor Ported the sync code to the latest libanki structure. Key points: No summary: The old style got each side to fetch ids+mod times and required the client to diff them and then request or bundle up the appropriate objects. Instead, we now get each side to send all changed objects, and it's the responsibility of the other side to decide what needs to be merged and what needs to be discarded. This allows us to skip a separate summary step, which saves scanning tables twice, and allows us to reduce server requests from 4 to 3. Schema changes: Certain operations that are difficult to merge (such as changing the number of fields in a model, or deleting models or groups) result in a full sync. The user is warned about it in the GUI before such schema-changing operations execute. Sync size: For now, we don't try to deal with large incremental syncs. Because the cards, facts and revlog can be large in memory (hundreds of megabytes in some cases), they would have to be chunked for the benefit of devices with a low amount of memory. Currently findChanges() uses the full fact/card objects which we're planning to send to the server. It could be rewritten to fetch a summary (just the id, mod & rep columns) which would save some memory, and then compare against blocks of a few hundred remote objects at a time. However, it's a bit more complicated than that: - If the local summary is huge it could exceed memory limits. Without a local summary we'd have to query the db for each record, which could be a lot slower. - We currently accumulate a list of remote records we need to add locally. This list also has the potential to get too big. We would need to periodically commit the changes as we accumulate them. - Merging a large amount of changes is also potentially slow on mobile devices. Given the fact that certain schema-changing operations require a full sync anyway, I think it's probably best to concentrate on a chunked full sync for now instead, as provided the user syncs periodically it should not be easy to hit the full sync limits except after bulk editing operations. Chunked partial syncing should be possible to add in the future without any changes to the deck format. Still to do: - deck conf merging - full syncing - new http proxy	2011-09-08 12:50:42 +09:00
Damien Elmes	be5c5a2018	move tags into deck; code into separate file - moved tags into json like previous changes, and dropped the unnecessary id - added tags.py for a tag manager - moved the tag utilities from utils into tags.py	2011-08-28 13:44:29 +09:00
Damien Elmes	2dfdfad6f2	update license link	2011-04-28 09:24:01 +09:00
Damien Elmes	8fcc6b3085	gpl3->agpl	2011-04-28 09:24:01 +09:00
Damien Elmes	511d6e89a1	remove progress handling code; we'll do it in the GUI or provide cb	2011-04-28 09:23:55 +09:00
Damien Elmes	2f27133705	drop sqlalchemy; massive refactor SQLAlchemy is a great tool, but it wasn't a great fit for Anki: - We often had to drop down to raw SQL for performance reasons. - The DB cursors and results were wrapped, which incurred a sizable performance hit due to introspection. Operations like fetching 50k records from a hot cache were taking more than twice as long to complete. - We take advantage of sqlite-specific features, so SQL language abstraction is useless to us. - The anki schema is quite small, so manually saving and loading objects is not a big burden. In the process of porting to DBAPI, I've refactored the database schema: - App configuration data that we don't need in joins or bulk updates has been moved into JSON objects. This simplifies serializing, and means we won't need DB schema changes to store extra options in the future. This change obsoletes the deckVars table. - Renamed tables: -- fieldModels -> fields -- cardModels -> templates -- fields -> fdata - a number of attribute names have been shortened Classes like Card, Fact & Model remain. They maintain a reference to the deck. To write their state to the DB, call .flush(). Objects no longer have their modification time manually updated. Instead, the modification time is updated when they are flushed. This also applies to the deck. Decks will now save on close, because various operations that were done at deck load will be moved into deck close instead. Operations like undoing buried card are cheap on a hot cache, but expensive on startup. Programmatically you can call .close(save=False) to avoid a save and a modification bump. This will be useful for generating due counts. Because of the new saving behaviour, the save and save as options will be removed from the GUI in the future. The q/a cache and field cache generating has been centralized. Facts will automatically rebuild the cache on flush; models can do so with model.updateCache(). Media handling has also been reworked. It has moved into a MediaRegistry object, which the deck holds. Refcounting has been dropped - it meant we had to compare old and new value every time facts or models were changed, and existed for the sole purpose of not showing errors on a missing media download. Instead we just media.registerText(q+a) when it's updated. The download function will be expanded to ask the user if they want to continue after a certain number of files have failed to download, which should be an adequate alternative. And we now add the file into the media DB when it's copied to th emedia directory, not when the card is commited. This fixes duplicates a user would get if they added the same media to a card twice without adding the card. The old DeckStorage object had its upgrade code split in a previous commit; the opening and upgrading code has been merged back together, and put in a separate storage.py file. The correct way to open a deck now is import anki; d = anki.Deck(path). deck.getCard() -> deck.sched.getCard() same with answerCard deck.getCard(id) returns a Card object now. And the DB wrapper has had a few changes: - sql statements are a more standard DBAPI: - statement() -> execute() - statements() -> executemany() - called like execute(sql, 1, 2, 3) or execute(sql, a=1, b=2, c=3) - column0 -> list	2011-04-28 09:23:53 +09:00
Damien Elmes	2613143fe9	improve dynamic indices, implement new queue	2011-04-28 09:23:28 +09:00
Damien Elmes	f828393de3	rename deck.s to a more understable deck.db; keep s for compat	2011-04-28 09:21:07 +09:00
Damien Elmes	9421a037f6	remove self explanatory module docstrings; strip trailing whitespace	2011-04-28 09:21:07 +09:00
Damien Elmes	472b68b831	don't backup when importing / saving as	2010-02-20 10:03:39 +09:00
Damien Elmes	3d81181323	bulk media support -> local media copy, always send media table	2009-06-19 11:50:31 +09:00
Damien Elmes	e6b207f7af	force media sync to go in one direction	2009-06-18 06:48:41 +09:00
Damien Elmes	91afe651b3	randomize after .anki import	2009-06-04 06:59:07 +09:00
Damien Elmes	95b8d655e6	remove shared cache mode, it's not needed	2009-03-18 22:45:43 +09:00
Damien Elmes	77a6488f6d	when importing, tag cards as new, add unit test	2009-02-14 03:03:42 +09:00
Damien Elmes	97359df499	add _ to anki10	2009-01-17 14:12:05 +09:00
Damien Elmes	334d126237	recording & noise profile support on linux	2009-01-17 01:05:39 +09:00
Damien Elmes	1fa7466dd9	progress for importing	2009-01-16 20:22:46 +09:00
Damien Elmes	75a61a00cc	remove card tags	2008-11-28 14:40:27 +09:00
Damien Elmes	c0e5bed6a6	sync sources, support media syncing in import/export again	2008-10-18 20:20:43 +09:00
Damien Elmes	5da3a0f5d3	initial commit from hg	2008-09-27 23:50:03 +09:00

22 commits