GitHub/Anki - OBJNULLs Forgejo: Beyond coding. We Forge.

mirror of https://github.com/ankitects/anki.git synced 2025-11-10 14:47:12 -05:00

Author	SHA1	Message	Date
RumovZ	5f9451f547	Add apkg import/export on backend (#1743 ) * Add apkg export on backend * Filter out missing media-paths at write time * Make TagMatcher::new() infallible * Gather export data instead of copying directly * Revert changes to rslib/src/tags/ * Reuse filename_is_safe/check_filename_safe() * Accept func to produce MediaIter in export_apkg() * Only store file folder once in MediaIter * Use temporary tables for gathering export_apkg() now accepts a search instead of a deck id. Decks are gathered according to the matched notes' cards. * Use schedule_as_new() to reset cards * ExportData → ExchangeData * Ignore ascii case when filtering system tags * search_notes_cards_into_table → search_cards_of_notes_into_table * Start on apkg importing on backend * Fix due dates in days for apkg export * Refactor import-export/package - Move media and meta code into appropriate modules. - Normalize/check for normalization when deserializing media entries. * Add SafeMediaEntry for deserialized MediaEntries * Prepare media based on checksums - Ensure all existing media files are hashed. - Hash incoming files during preparation to detect conflicts. - Uniquify names of conflicting files with hash (not notetype id). - Mark media files as used while importing notes. - Finally copy used media. * Handle encoding in `replace_media_refs()` * Add trait to keep down cow boilerplate * Add notetypes immediately instaed of preparing * Move target_col into Context * Add notes immediately instaed of preparing * Note id, not guid of conflicting notes * Add import_decks() * decks_configs → deck_configs * Add import_deck_configs() * Add import_cards(), import_revlog() * Use dyn instead of generic for media_fn Otherwise, would have to pass None with type annotation in the default case. * Fix signature of import_apkg() * Fix search_cards_of_notes_into_table() * Test new functions in text.rs * Add roundtrip test for apkg (stub) * Keep source id of imported cards (or skip) * Keep source ids of imported revlog (or skip) * Try to keep source ids of imported notes * Make adding notetype with id undoable * Wrap apkg import in transaction * Keep source ids of imported deck configs (or skip) * Handle card due dates and original due/did * Fix importing cards/revlog Card ids are manually uniquified. * Factor out card importing * Refactor card and revlog importing * Factor out card importing Also handle missing parents . * Factor out note importing * Factor out media importing * Maybe upgrade scheduler of apkg * Fix parent deck gathering * Unconditionally import static media * Fix deck importing edge cases Test those edge cases, and add some global test helpers. * Test note importing * Let import_apkg() take a progress func * Expand roundtrip apkg test * Use fat pointer to avoid propogating generics * Fix progress_fn type * Expose apkg export/import on backend * Return note log when importing apkg * Fix archived collection name on apkg import * Add CollectionOpWithBackendProgress * Fix wrong Interrupted Exception being checked * Add ClosedCollectionOp * Add note ids to log and strip HTML * Update progress when checking incoming media too * Conditionally enable new importing in GUI * Fix all_checksums() for media import Entries of deleted files are nulled, not removed. * Make apkg exporting on backend abortable * Return number of notes imported from apkg * Fix exception printing for QueryOp as well * Add QueryOpWithBackendProgress Also support backend exporting progress. * Expose new apkg and colpkg exporting * Open transaction in insert_data() Was slowing down exporting by several orders of magnitude. * Handle zstd-compressed apkg * Add legacy arg to ExportAnkiPackage Currently not exposed on the frontend * Remove unused import in proto file * Add symlink for typechecking of import_export_pb2 * Avoid kwargs in pb message creation, so typechecking is not lost Protobuf's behaviour is rather subtle and I had to dig through the docs to figure it out: set a field on a submessage to automatically assign the submessage to the parent, or call SetInParent() to persist a default version of the field you specified. * Avoid re-exporting protobuf msgs we only use internally * Stop after one test failure mypy often fails much faster than pylint * Avoid an extra allocation when extracting media checksums * Update progress after prepare_media() finishes Otherwise the bulk of the import ends up being shown as "Checked: 0" in the progress window. * Show progress of note imports Note import is the slowest part, so showing progress here makes the UI feel more responsive. * Reset filtered decks at import time Before this change, filtered decks exported with scheduling remained filtered on import, and maybe_remove_from_filtered_deck() moved cards into them as their home deck, leading to errors during review. We may still want to provide a way to preserve filtered decks on import, but to do that we'll need to ensure we don't rewrite the home decks of cards, and we'll need to ensure the home decks are included as part of the import (or give an error if they're not). https://github.com/ankitects/anki/pull/1743/files#r839346423 * Fix a corner-case where due dates were shifted by a day This issue existed in the old Python code as well. We need to include the user's UTC offset in the exported file, or days_elapsed falls back on the v1 cutoff calculation, which may be a day earlier or later than the v2 calculation. * Log conflicting note in remapped nt case * take_fields() → into_fields() * Alias `[u8; 20]` with `Sha1Hash` * Truncate logged fields * Rework apkg note import tests - Use macros for more helpful errors. - Split monolith into unit tests. - Fix some unknown error with the previous test along the way. (Was failing after 969484de4388d225c9f17d94534b3ba0094c3568.) * Fix sorting of imported decks Also adjust the test, so it fails without the patch. It was only passing before, because the parent deck happened to come before the inconsistently capitalised child alphabetically. But we want all parent decks to be imported before their child decks, so their children can adopt their capitalisation. * target[_id]s → existing_card[_id]s * export_collection_extracting_media() → ... export_into_collection_file() * target_already_exists→card_ordinal_already_exists * Add search_cards_of_notes_into_table.sql * Imrove type of apkg export selector/limit * Remove redundant call to mod_schema() * Parent tooltips to mw * Fix a crash when truncating note text String::truncate() is a bit of a footgun, and I've hit this before too :-) * Remove ExportLimit in favour of separate classes * Remove OpWithBackendProgress and ClosedCollectionOp Backend progress logic is now in ProgressManager. QueryOp can be used for running on closed collection. Also fix aborting of colpkg exports, which slipped through in #1817. * Tidy up import log * Avoid QDialog.exec() * Default to excluding scheuling for deck list deck * Use IncrementalProgress in whole import_export code * Compare checksums when importing colpkgs * Avoid registering changes if hashes are not needed * ImportProgress::Collection → ImportProgress::File * Make downgrading apkgs depend on meta version * Generalise IncrementableProgress And use it in entire import_export code instead. * Fix type complexity lint * Take count_map for IncrementableProgress::get_inner * Replace import/export env with Shift click * Accept all args from update() for backend progress * Pass fields of ProgressUpdate explicitly * Move update_interval into IncrementableProgress * Outsource incrementing into Incrementor * Mutate ProgressUpdate in progress_update callback * Switch import/export legacy toggle to profile setting Shift would have been nice, but the existing shortcuts complicate things. If the user triggers an import with ctrl+shift+i, shift is unlikely to have been released by the time our code runs, meaning the user accidentally triggers the new code. We could potentially wait a while before bringing up the dialog, but then we're forced to guess at how long it will take the user to release the key. One alternative would be to use alt instead of shift, but then we need to trigger our shortcut when that key is pressed as well, and it could potentially cause a conflict with an add-on that already uses that combination. * Show extension in export dialog * Continue to provide separate options for schema 11+18 colpkg export * Default to colpkg export when using File>Export * Improve appearance of combo boxes when switching between apkg/colpkg + Deal with long deck names * Convert newlines to spaces when showing fields from import Ensures each imported note appears on a separate line * Don't separate total note count from the other summary lines This may come down to personal preference, but I feel the other counts are equally as important, and separating them feels like it makes it a bit easier to ignore them. * Fix 'deck not normal' error when importing a filtered deck for the 2nd time * Fix [Identical] being shown on first import * Revert "Continue to provide separate options for schema 11+18 colpkg export" This reverts commit `8f0b2c175f`. Will use a different approach * Move legacy support into a separate exporter option; add to apkg export * Adjust 'too new' message to also apply to .apkg import case * Show a better message when attempting to import new apkg into old code Previously the user could end seeing a message like: UnicodeDecodeError: 'utf-8' codec can't decode byte 0xb5 in position 1: invalid start byte Unfortunately we can't retroactively fix this for older clients. * Hide legacy support option in older exporting screen * Reflect change from paths to fnames in type & name * Make imported decks normal at once Then skip special casing in update_deck(). Also skip updating description if new one is empty. Co-authored-by: Damien Elmes <gpg@ankiweb.net>	2022-05-02 21:12:46 +10:00
RumovZ	872b6df22a	Optimise searching in (all) fields (#1622 ) * Avoid rebuilding regex in field search * Special case search in all fields * Don't repeat mid nodes in field search sql Small speed gain for searches like `:re:foo` and reduces the sql tree depth if a lot of field names of the same notetype match. Add sql function to match fields with regex * Optimise used field search algorithm - Searching in all fields is a special case. - Using native SQL comparison is preferred. - For Regex, use newly added SQL function. * Please clippy * Avoid pyramid of doom * nt_fields -> matched_fields * Add tests for regex and all field searches * minor tweaks for readability (dae)	2022-01-24 20:30:08 +10:00
RumovZ	3e0c9dc866	New TTS/AV tag handling (#1559 ) * Add new `card_rendering` mod Parses a text with av/tts tags and strips or extracts tags. * Replace old `extract_av_tags` and `strip_av_tags` ... with new `card_rendering` mod * ressource -> resource * Add AV prettifier for use in browser table * Accept String in av tag routines ... and avoid redundant writes if no changes need to be made. * add benchmarking with criterion; make links test optional (dae) cargo install cargo-criterion, then run ./bench.sh * performance comparison: creating HashMap up front (dae) the previous solution: anki_tag_parse time: [1.8401 us 1.8437 us 1.8476 us] this solution: anki_tag_parse time: [2.2420 us 2.2447 us 2.2477 us] change: [+21.477% +21.770% +22.066%] (p = 0.00 < 0.05) Performance has regressed. * Revert "performance comparison: creating HashMap up front" (dae) This reverts commit `f19126a2f1`. * add missing header * Write error message if tts lang is missing * `Tag` -> `Directive`	2021-12-17 19:04:42 +10:00
RumovZ	3ab9712c18	Stop trimming filename references before encoding (#1462 ) Closes #1430	2021-10-28 19:22:51 +10:00
Damien Elmes	53fe7e574e	handle ampersand entities in image filenames In the old HTML editor, filenames were % escaped before feeding them to beautifulsoup, causing bare ampersands to be left alone. The new HTML editor reads content from the DOM, where a bare ampersand has been transformed into an &, and that gets saved back into the field, so the media check now needs to deal with it for images as well. https://forums.ankiweb.net/t/causing-problems-with-image-names/12171	2021-08-19 23:43:40 +10:00
Damien Elmes	2f434dd74d	fix comment + copy/paste error	2021-07-17 09:02:14 +10:00
Damien Elmes	bf507cca98	move from Python's URI escaping to IRI escaping in Rust Should make non-Latin text readable in the HTML editor, without the breakages reverted in the previous change.	2021-07-16 10:38:00 +10:00
Damien Elmes	b392020798	fix clippy lints for latest Rust	2021-06-21 13:09:36 +10:00
Henrik Giesel	0f658de702	Add escape_anki_wildcards_for_search_node	2021-06-16 09:25:27 +02:00
Henrik Giesel	c5faf39d7c	Make Browser root nodes use "_*" uniformly	2021-06-16 17:19:21 +10:00
Damien Elmes	220fca9a1d	handle <br/> when rendering a single line + case-insensitive matching https://forums.ankiweb.net/t/html-editor-modifies-note-when-a-field-with-break-tags-is-opened/10772	2021-06-14 13:05:48 +10:00
Damien Elmes	64ebc32b3d	tidy up Rust imports rustfmt can do this automatically, but only when run with a nightly toolchain, so it needs to be manually done for now - see rslib/rusfmt.toml	2021-04-18 18:38:54 +10:00
Damien Elmes	dc81a7fed0	use mixed case for abbreviations in Rust code So, this is fun. Apparently "DeckId" is considered preferable to the "DeckID" were were using until now, and the latest clippy will start warning about it. We could of course disable the warning, but probably better to bite the bullet and switch to the naming that's generally considered best.	2021-03-27 19:53:33 +10:00
Shaun Ren	1f3751d191	Fix extraneous whitespaces from strip_html_for_tts	2021-03-25 11:44:42 -04:00
RumovZ	9151bfb53e	Return input if decode_entities() encounters error	2021-03-22 12:08:22 +01:00
RumovZ	e931a429b3	Add html_to_text_line() on backend	2021-03-20 12:00:45 +01:00
Damien Elmes	ab790c1d14	initial work on moving v2 card answering into backend Not plugged into the Python code yet. Still a work in progress. Other changes: - move a bunch of From implementations out of the giant backend/mod.rs file into separate submodules. - reorder backend methods to match proto order - fix some clippy lints	2021-02-20 14:48:07 +10:00
Damien Elmes	242b4ea505	switch search parser to using owned values I was a bit too enthusiastic with using borrowed values in structs earlier on in the Rust porting. In this case any performance gains are dwarfed by the cost of querying the DB, and using owned values here simplifies the code, and will make it easier to parse a fragment in the From<SearchTerm> impl.	2021-02-11 12:19:36 +10:00
Damien Elmes	ded626f0b9	render deck description with markdown; strip images To support images on that screen, we'll first need to adjust the base url for each platform, or rewrite the local image URLs, as otherwise they are resolved to _anki/pages/...	2021-02-06 15:02:40 +10:00
abdo	72e8f9d640	Merge branch 'master' of https://github.com/ankitects/anki into tagtree	2021-01-12 23:31:58 +03:00
abdo	b276ce3dd5	Hierarchical tags	2021-01-09 17:10:13 +03:00
RumovZ	9ef691c06f	Provide filter searches through backend	2021-01-09 10:50:08 +01:00
abdo	197d665de8	Fix duplicate check not decoding entities This is a regression introduced in `358d0f957e` See https://forums.ankiweb.net/t/bug-duplicates-not-detecting-on-paste/5753	2020-12-14 15:13:00 +03:00
Damien Elmes	3923f56cbf	fix clippy lints	2020-11-24 20:13:05 +10:00
RumovZ	347c547e10	Add tests for conversion functions in text.rs	2020-11-20 09:45:53 +01:00
RumovZ	0fc84d19b2	Replace text.rs/text_to_re with text.rs/to_re	2020-11-20 09:23:25 +01:00
RumovZ	785540bddc	Revert changes to normalisation handling Handle norm calls individually in write_search_node_to_sql again.	2020-11-18 23:46:27 +01:00
RumovZ	c185fb966b	Merge branch 'master' into rework-search-parser Conflicts: rslib/src/search/sqlwriter.rs	2020-11-18 09:04:04 +01:00
RumovZ	91873d68eb	Fix RE in to_custom_re of text.rs Match every single (potentially escaped) character of the string, so they can be escaped properly.	2020-11-17 15:39:54 +01:00
RumovZ	8c02c6e205	Split unescaping between parser and writer * Unescape wildcards in writer instead of parser. * Move text conversion functions to text.rs. * Implicitly norm when converting text. * Revert to using collection when comparing tags but add escape support.	2020-11-17 12:49:37 +01:00
RumovZ	0cff65e5a8	Fix bugs and inconsistencies in the search parser	2020-11-12 17:27:50 +01:00
Andreas Reis	e68a40f13e	cleanup / renames ・ soundRegexps → sound_regexps ・ htmlRegexps → html_media_regexps ・ HTML_TAGS → HTML_MEDIA_TAGS ・ escapeImages → escape_media_filenames + alias ・ strip_html_preserving_image_filenames → strip_html_preserving_media_filenames	2020-11-10 14:53:04 +01:00
Andreas Reis	6e9aaad11e	Add audio & object tags to media check Makes the media check recognize files in <audio> and <object> tags as used. They've been observed/supported by the WebView (checked: Anki, AnkiDroid) since just about forever already and are extremely useful if one knows a thing about web dev.	2020-10-25 13:09:57 +01:00
Damien Elmes	c82a084edf	handle quoted html chars in media check https://forums.ankiweb.net/t/unable-to-play-longer-audio-on-cards/1313/30	2020-09-04 09:36:38 +10:00
Damien Elmes	9fcd6c66f4	fix nonbreaking spaces breaking media https://forums.ankiweb.net/t/unable-to-play-longer-audio-on-cards/1313	2020-08-30 11:23:12 +10:00
Damien Elmes	9e53c84a35	fix globs not working in bulk tag add/remove	2020-08-17 18:14:00 +10:00
Damien Elmes	805a3a710e	split note types into separate tables - store the config in protobuf instead of json - still loading+saving in bulk for now - code using the schema11 structs needs to be migrated	2020-05-12 21:13:33 +10:00
Damien Elmes	66809dd8a3	ignore empty sound tags https://github.com/ankitects/anki/pull/612	2020-05-12 20:53:50 +10:00
Damien Elmes	c95983ac1f	preserve entities when stripping HTML for MathJax https://anki.tenderapp.com/discussions/ankidesktop/40987-how-to-render-angled-brackets	2020-04-30 11:17:38 +10:00
Damien Elmes	51a379de23	add search that ignores combining chars On a test of a ~40k card collection, the 'ignore accents' add-on takes about 1150ms, and this code takes about 70ms.	2020-03-21 15:15:59 +10:00
Damien Elmes	08e64d246d	don't require wildcard for unicode case folding in search	2020-03-21 12:44:56 +10:00
Damien Elmes	9f3cc0982d	deck searching A bit more complicated than it needs to be, as we don't have the full deck manager infrastructure yet.	2020-03-20 21:15:23 +10:00
Damien Elmes	3a1fc74ec3	remove some unused imports	2020-02-29 15:21:11 +10:00
Damien Elmes	ee27711b65	remove redundant test_ prefix	2020-02-17 08:40:17 +10:00
Damien Elmes	c890ef871e	include LaTeX png/svg files when checking for unused media	2020-02-17 08:40:17 +10:00
Damien Elmes	fabfcb0338	gather field references in Rust; media check now mostly complete	2020-02-17 08:40:17 +10:00
Damien Elmes	f7c26724f3	nfc helper	2020-02-17 08:40:17 +10:00
Damien Elmes	c075191697	reuse reveal_cloze_text() for LaTeX cloze expansion	2020-01-28 07:40:44 +10:00
Damien Elmes	9ad80f4d2c	move cloze-related code into a separate file	2020-01-27 20:41:23 +10:00
Damien Elmes	21cbb5a766	support speed control in tts tags	2020-01-26 14:31:07 +10:00

1 2

55 commits