erinhmclark
077b56c150
Merge GSheet Feeder and Database.
2025-03-04 14:05:19 +00:00
Patrick Robertson
f54d6519a8
Fix sorting of steps in the outputted file
2025-03-04 11:51:26 +00:00
Patrick Robertson
07ee773a54
Better drag & drop + keep comments in file
2025-03-04 10:54:16 +00:00
erinhmclark
a705a78632
Fix instagram_extractor.py typo in config value.
2025-03-03 21:06:09 +00:00
Patrick Robertson
dcaf7639be
Rename 'upgrading...' page to 'upgrading from...' because it's also valid for 0.13+ versions
2025-03-03 19:58:47 +00:00
Patrick Robertson
0b5a0fcb32
Better error logs if users have XXXX_archiver modules enabled in config
2025-03-03 19:57:09 +00:00
Patrick Robertson
1fe023cd70
Throw a nicer error if a user has an orchestration.yaml file in the old format (feeder: / archivers: / formatter: )
2025-03-03 19:51:55 +00:00
Patrick Robertson
a47e18ef9a
Bump gecko driver to 0.36.0
2025-03-03 16:00:11 +00:00
Patrick Robertson
0dfab2d1bc
Add some code to attempt to click the cookies banners on various websites
2025-03-03 15:55:04 +00:00
Patrick Robertson
dea0a49600
Download correct gecko-driver for the platform + fix setting executable path when running in Docker
...
Fixes #232
2025-03-03 15:41:44 +00:00
Erin Clark
011ded2bde
Merge pull request #225 from bellingcat/small_issues
...
## GSheets Columns updates
- Update the available columns in the Google Sheet Feeder and Database.
- Update the Sheet Template to reflect this.
## Other Fixes
- Ensure test file cleanup.
- Additional tests.
- Correctly mark download test.
- Small typos.
2025-03-03 13:06:27 +00:00
Patrick Robertson
a88a37d0a5
Hook in to RTD theme to set react theme
2025-03-03 11:56:23 +00:00
Patrick Robertson
a0869bb3b2
Fixed up timestamp verifying - waiting on issue with rfc-client to be fixed
...
Ref: https://github.com/trailofbits/rfc3161-client/issues/104#issuecomment-2693890607
2025-03-03 10:28:30 +00:00
Patrick Robertson
9845804277
Fix up TODO plus add comments on integration into RTD page
2025-03-03 09:18:19 +00:00
Patrick Robertson
cc14e5cb9f
Remove extra html/head tag from page - now it's embedded in RTD
2025-03-03 09:06:40 +00:00
Patrick Robertson
6ba79049d9
Capitalize help text
2025-02-27 22:16:33 +00:00
Patrick Robertson
7620a671d1
Overwrite settings_base file
2025-02-27 22:02:44 +00:00
Patrick Robertson
54a2a19dd7
Also build auto-archiver
2025-02-27 21:56:36 +00:00
Patrick Robertson
3eb4ab41b8
Also generate the schema on each run
2025-02-27 21:38:55 +00:00
Patrick Robertson
65a9885d86
A few more manifest types
2025-02-27 21:33:04 +00:00
Patrick Robertson
4ee1e75aa2
Fix readthedocs config file
2025-02-27 21:24:34 +00:00
Patrick Robertson
1141c00e9a
Remove unused files, set up for RTD
2025-02-27 21:23:38 +00:00
Patrick Robertson
15da907e81
Add a bit of typescripting
2025-02-27 15:58:30 +00:00
Patrick Robertson
2ec44f4170
Documentation on building the settings page
2025-02-27 15:42:37 +00:00
Patrick Robertson
1e92c03b1d
Tweaks to settings page + more declarations in manifests
2025-02-27 15:21:11 +00:00
Patrick Robertson
efe9fdf915
Tidy ups to config editor page
2025-02-27 13:02:50 +00:00
erinhmclark
4280791f07
Fix mocking in test_wayback_enricher.py.
2025-02-27 11:25:58 +00:00
Patrick Robertson
f58f110436
Check at least 1 URL provided for new cli_feeder module rewrite
2025-02-26 17:59:13 +00:00
Patrick Robertson
70d89c71ce
Fully-working settings page editor
2025-02-26 17:02:49 +00:00
Patrick Robertson
bb961b131c
Turn cli_feeder *back* into a module, it's better like this for settings etc, documentation etc.
2025-02-26 15:41:33 +00:00
Patrick Robertson
e467fc90c2
Merge branch 'main' into settings_page
2025-02-26 15:32:07 +00:00
erinhmclark
8124bb831d
Merge branch 'main' into small_issues
...
# Conflicts:
# src/auto_archiver/core/base_module.py
# src/auto_archiver/utils/misc.py
2025-02-26 13:19:49 +00:00
erinhmclark
b2e654aef9
Remove context manager from test_pdq_hash_enricher.py
2025-02-26 12:57:33 +00:00
erinhmclark
9157846930
Add docstrings to explain date formats.
2025-02-26 10:01:52 +00:00
Patrick Robertson
600f43e790
Set up structure for react
2025-02-26 09:34:44 +00:00
Patrick Robertson
afc117a229
Get downloading certs working
2025-02-26 09:33:56 +00:00
erinhmclark
696aafb52d
Update gsheet_feeder references in tests.
2025-02-25 21:38:41 +00:00
erinhmclark
75380b0716
Merge GSheet Feeder and Database.
2025-02-25 21:32:32 +00:00
erinhmclark
35b5ab2eb1
Update poetry.lock
2025-02-25 20:17:48 +00:00
erinhmclark
83a08dd215
Update date parsing to use dateutil.parser in misc.py
2025-02-25 20:17:31 +00:00
erinhmclark
9bc6dd5c3c
Add set_content into generic_extractor.py.
2025-02-25 20:07:00 +00:00
erinhmclark
cf1219f798
Add text content into gsheet.
2025-02-25 20:06:44 +00:00
Patrick Robertson
4dcb77c29f
Merge branch 'main' into timestamping_rewrite
2025-02-25 17:10:55 +00:00
Patrick Robertson
1ad158c016
Merge pull request #211 from bellingcat/docs_improvements
...
Docs tidyups, howto on logging and authentication, remove exit(), small fixes
2025-02-25 14:13:13 +00:00
erinhmclark
1df5129268
Small typos.
2025-02-25 14:08:38 +00:00
erinhmclark
73b434aafc
Tests for test_vk_extractor.py.
2025-02-25 14:08:28 +00:00
erinhmclark
2d276cb9c4
Fix tmp test file.
2025-02-25 14:08:14 +00:00
Patrick Robertson
898faf6fe4
Further WIP - currently working on verify_signed
2025-02-25 12:08:08 +00:00
Patrick Robertson
6987a4827e
Set poetry packages - remove tsp_client and update cryptography
2025-02-25 11:57:20 +00:00
Patrick Robertson
f8e846d59a
Create facebook dropin - working for images + text. CAVEAT: only gets the first ~100 chars of the post at the moment
2025-02-25 11:44:35 +00:00