erinhmclark
|
d9d936c2ca
|
Thumbnail enricher fix seconds to minutes.
|
2025-02-12 12:22:27 +00:00 |
|
Patrick Robertson
|
460a71649c
|
Merge pull request #190 from bellingcat/docs_update
Docs improvement
|
2025-02-12 12:38:04 +01:00 |
|
Patrick Robertson
|
5b481f72ab
|
Tidy ups to manifests for docs
|
2025-02-11 20:03:53 +00:00 |
|
Patrick Robertson
|
4c119b4db8
|
Add missing manifest for atlos_storage
|
2025-02-11 20:03:45 +00:00 |
|
Patrick Robertson
|
1ee7981c6e
|
Add YAML config to the module docs
|
2025-02-11 19:42:03 +00:00 |
|
Patrick Robertson
|
1b976f4c09
|
Remove unused atlos util functions
|
2025-02-11 18:49:54 +00:00 |
|
Patrick Robertson
|
3787577a96
|
Screenshot enricher depends on geckodriver not chromedriver
|
2025-02-11 18:18:52 +00:00 |
|
Patrick Robertson
|
d0c379a3ba
|
WIP - timestamping enricher
|
2025-02-11 18:18:19 +00:00 |
|
Patrick Robertson
|
ea728a7a97
|
TODO on facebook dropin not working
|
2025-02-11 15:56:12 +00:00 |
|
msramalho
|
91f1ebf7b3
|
fix temp for yandex new shortlink
|
2025-02-11 15:23:16 +00:00 |
|
msramalho
|
c720541de2
|
merge conflicts
|
2025-02-11 15:22:06 +00:00 |
|
Patrick Robertson
|
7bb4d68a22
|
Merge branch 'load_modules' into timestamping_rewrite
|
2025-02-11 15:21:31 +00:00 |
|
msramalho
|
e507fc81d2
|
improves mimetype guessing, previously file.sub.something would not have an extension
|
2025-02-11 15:02:49 +00:00 |
|
msramalho
|
5478ed3860
|
bsky fix media fetching
|
2025-02-11 15:02:00 +00:00 |
|
msramalho
|
47d1dc9d47
|
typing warnings fixed
|
2025-02-11 15:01:37 +00:00 |
|
Patrick Robertson
|
2d87935042
|
Start on opentimestamps enricher
|
2025-02-11 14:54:46 +00:00 |
|
Patrick Robertson
|
29901da601
|
Merge branch 'load_modules' into docs_update
|
2025-02-11 14:10:56 +00:00 |
|
Patrick Robertson
|
7d87b858d6
|
Merge branch 'load_modules' into docs_update
|
2025-02-11 13:09:38 +00:00 |
|
erinhmclark
|
c8cd7ea63c
|
Merge branch 'load_modules' into add_module_tests
# Conflicts:
# src/auto_archiver/modules/telethon_extractor/telethon_extractor.py
|
2025-02-11 13:08:08 +00:00 |
|
msramalho
|
977618b4ce
|
doc: adds note about telethon vs telegram extractors
|
2025-02-11 13:04:59 +00:00 |
|
msramalho
|
d90d3cec28
|
fix telethon_extractor setup
|
2025-02-11 13:03:18 +00:00 |
|
msramalho
|
977f06c37a
|
renames api_db property for clarity
|
2025-02-11 12:56:33 +00:00 |
|
msramalho
|
5c59029221
|
updates api_db for new API endpoint
|
2025-02-11 12:53:58 +00:00 |
|
msramalho
|
4eeb39477c
|
improves gsheetdb feedback on retrieve sheet failure
|
2025-02-11 12:53:46 +00:00 |
|
msramalho
|
6fdd5f0e66
|
fix cases of single : vs :: in entrypoint
|
2025-02-11 12:53:12 +00:00 |
|
Patrick Robertson
|
2650cd8fb2
|
Use a script to auto-generate documentation for the core modules from the manifest file
|
2025-02-10 22:51:04 +00:00 |
|
erinhmclark
|
8d894066f2
|
Merge branch 'load_modules' into add_module_tests
# Conflicts:
# src/auto_archiver/modules/gsheet_feeder/gsheet_feeder.py
# src/auto_archiver/utils/misc.py
|
2025-02-10 19:00:05 +00:00 |
|
erinhmclark
|
3dae2337a1
|
remove cdn_url check before storage.
|
2025-02-10 18:56:46 +00:00 |
|
erinhmclark
|
e97ccf8a73
|
Separate setup() and module_setup().
|
2025-02-10 18:07:47 +00:00 |
|
erinhmclark
|
2c3d1f591f
|
Separate setup() and module_setup().
|
2025-02-10 17:25:15 +00:00 |
|
msramalho
|
12f14cccc9
|
fixes gsheet feeder<->db connection via context.
|
2025-02-10 16:58:35 +00:00 |
|
msramalho
|
ab6cf52533
|
fixes bad hash initialization
|
2025-02-10 16:45:28 +00:00 |
|
erinhmclark
|
c4bb667cec
|
Merge branch 'load_modules' into add_module_tests
# Conflicts:
# src/auto_archiver/modules/s3_storage/s3_storage.py
# src/auto_archiver/utils/gsheet.py
# src/auto_archiver/utils/misc.py
|
2025-02-10 16:17:08 +00:00 |
|
erinhmclark
|
f311621e58
|
Small fixes.
Add timestamp helper method.
|
2025-02-10 15:57:42 +00:00 |
|
msramalho
|
15abf686b1
|
decouples s3_storage from hash_enricher
|
2025-02-10 15:48:54 +00:00 |
|
msramalho
|
8fb3dc754b
|
fixing telethon extractor to use default entrypoint
|
2025-02-10 14:59:51 +00:00 |
|
Patrick Robertson
|
63aba6ad39
|
Fix sphinx-autoapi imports
|
2025-02-07 21:54:49 +01:00 |
|
erinhmclark
|
950624dd4b
|
Fix S3 storage to media in whisper_enricher.py.
|
2025-02-07 20:26:00 +00:00 |
|
erinhmclark
|
2920cf685f
|
Small fixes to whisper_enricher.py.
|
2025-02-07 12:35:40 +00:00 |
|
erinhmclark
|
e9ad1e1b85
|
Pass media to storage cdn_call
|
2025-02-06 22:01:55 +00:00 |
|
erinhmclark
|
266c7a14e6
|
Context related fixes, some more tests.
|
2025-02-06 16:53:00 +00:00 |
|
erinhmclark
|
5b0bad832f
|
Updated test, test metadata
|
2025-02-06 10:11:56 +00:00 |
|
erinhmclark
|
52542812dc
|
Merge tests from version with context.
|
2025-02-05 16:42:58 +00:00 |
|
Patrick Robertson
|
91ca325fd5
|
Update yt-dlp to latest version + remove code no longer needed from bluesky dropin
|
2025-02-04 17:46:46 +01:00 |
|
Patrick Robertson
|
034197a81f
|
Fix typos in csv feeder docs (in manifest)
|
2025-02-04 13:40:07 +01:00 |
|
Patrick Robertson
|
78e6418249
|
Unit tests for csv feeder + fix some bugs
|
2025-02-04 13:37:26 +01:00 |
|
Patrick Robertson
|
b301f60ea3
|
Fix using validators set in __manifest__.py
E.g. you can use the validator 'is_file' to check if a config is a valid file
|
2025-02-04 13:37:26 +01:00 |
|
Patrick Robertson
|
c574b694ed
|
Set up screenshot enricher to use authentication/cookies
|
2025-02-03 17:25:59 +01:00 |
|
Patrick Robertson
|
7ec328ab40
|
Remove cookie options from generic_extractor - it now uses 'authentication' global settings :D
|
2025-02-03 16:04:36 +01:00 |
|
Patrick Robertson
|
7a2be5a0da
|
Add cookie extraction to 'authentication' options, get generic_extractor working using this info
|
2025-02-03 16:03:07 +01:00 |
|