forked-synapse

mirror of https://mau.dev/maunium/synapse.git synced 2024-10-01 01:36:05 -04:00

Author	SHA1	Message	Date
Mathieu Velten	dac97642e4	Implements admin API to lock an user (MSC3939) (#15870 )	2023-08-10 09:10:55 +00:00
Shay	0328b56468	Support MSC3814: Dehydrated Devices Part 2 (#16010 )	2023-08-08 12:04:46 -07:00
reivilibre	f3dc6dc19f	Remove old rows from the `cache_invalidation_stream_by_instance` table automatically. (This table is not used when Synapse is configured to use SQLite.) (#15868 ) * Add a cache invalidation clean-up task * Run the cache invalidation stream clean-up on the background worker * Tune down * call_later is in millis! * Newsfile Signed-off-by: Olivier Wilkinson (reivilibre) <oliverw@matrix.org> * fixup! Add a cache invalidation clean-up task * Update synapse/storage/databases/main/cache.py Co-authored-by: Eric Eastwood <erice@element.io> * Update synapse/storage/databases/main/cache.py Co-authored-by: Eric Eastwood <erice@element.io> * MILLISEC -> MS * Expand on comment * Move and tweak comment about Postgres * Use `wrap_as_background_process` --------- Signed-off-by: Olivier Wilkinson (reivilibre) <oliverw@matrix.org> Co-authored-by: Eric Eastwood <erice@element.io>	2023-08-08 11:10:07 +01:00
Patrick Cloke	d98a43d922	Stabilize support for MSC3970: updated transaction semantics (scope to `device_id`) (#15629 ) For now this maintains compatible with old Synapses by falling back to using transaction semantics on a per-access token. A future version of Synapse will drop support for this.	2023-08-04 07:47:18 -04:00
Erik Johnston	ae55cc1e6b	Add ability to wait for locks and add locks to purge history / room deletion (#15791 ) c.f. #13476	2023-07-31 10:58:03 +01:00
Anshul Madnawat	58f8305114	Inline SQL queries using boolean parameters (#15525 ) SQLite now supports TRUE and FALSE constants, simplify some queries by inlining those instead of passing them as arguments.	2023-07-26 18:45:47 +00:00
Mathieu Velten	8ebfd577e2	Bump DB version to 79 since synapse v1.88 was already there (#15998 )	2023-07-26 14:51:44 +02:00
Shay	f08d05dd2c	Actually stop reading from column `user_id` of tables `profiles` (#15955 )	2023-07-23 16:30:54 -07:00
Erik Johnston	fc1e534e41	Speed up updating state in large rooms (#15971 ) This should speed up updating state in rooms with lots of state.	2023-07-20 15:51:28 +01:00
Erik Johnston	19796e20aa	Fix bad merge of #15933 (#15958 ) This was because we reverted the bump of the schema version, so we were not applying the new deltas.	2023-07-19 12:17:08 +00:00
Erik Johnston	40a3583ba1	Fix race in triggers for read/write locks. (#15933 )	2023-07-19 12:06:38 +01:00
Shay	cb6e2c6cc7	Fix background schema updates failing over a large upgrade gap (#15887 )	2023-07-18 16:59:27 -07:00
Olivier Wilkinson (reivilibre)	8e8431bc6e	Merge branch 'master' into develop	2023-07-18 16:45:39 +01:00
Patrick Cloke	6d81aec09f	Support room version 11 (#15912 ) And fix a bug in the implementation of the updated redaction format (MSC2174) where the top-level redacts field was not properly added for backwards-compatibility.	2023-07-18 08:44:59 -04:00
Shay	e625c3dca0	Revert "Stop writing to column `user_id` of tables `profiles` and `user_filters`. (#15953 ) * Revert "Stop writing to column `user_id` of tables `profiles` and `user_filters` (#15787)" This reverts commit `f25b0f8808`. * newsfragement	2023-07-18 11:44:09 +01:00
Mathieu Velten	8eb7bb975e	Mark get_user_in_directory private since only used in tests (#15884 )	2023-07-12 11:09:13 +02:00
Erik Johnston	e55a9b3e41	Fix downgrading to previous version of Synapse (#15907 ) We do this by marking the constraint as deferrable.	2023-07-10 16:24:42 +01:00
Shay	f25b0f8808	Stop writing to column `user_id` of tables `profiles` and `user_filters` (#15787 )	2023-07-07 09:23:27 -07:00
Erik Johnston	39d131b016	Add basic read/write lock (#15782 )	2023-07-05 17:25:00 +01:00
Eric Eastwood	ce857c05d5	Add tracing to media `/upload` endpoint (#15850 ) Add tracing instrumentation to media `/upload` code paths to investigate https://github.com/matrix-org/synapse/issues/15841	2023-07-05 10:22:21 -05:00
Jason Little	4cf9f92f39	Fix could not serialize access due to concurrent `DELETE` from presence_stream (#15826 ) * Change update_presence to have a isolation level of READ_COMMITTED * changelog	2023-07-05 11:44:02 +01:00
Erik Johnston	95a96b21eb	Add foreign key constraint to `event_forward_extremities`. (#15751 )	2023-07-05 09:43:19 +00:00
Michael Weimann	c8e81898b6	Add not_user_type param to the list accounts admin API (#15844 ) Signed-off-by: Michael Weimann <michaelw@element.io>	2023-07-04 15:03:20 -07:00
pacien	07d7cbfe69	devices: use combined ANY clause for faster cleanup (#15861 ) Old device entries for the same user were being removed in individual SQL commands, making the batch take way longer than necessary. This combines the commands into a single one with a IN/ANY clause. Example of log entry before the change, regularly observed with "log_min_duration_statement = 10000" in PostgreSQL's config: LOG: duration: 42538.282 ms statement: DELETE FROM device_lists_stream WHERE user_id = '@someone' AND device_id = 'someid1' AND stream_id < 123456789 ; DELETE FROM device_lists_stream WHERE user_id = '@someone' AND device_id = 'someid2' AND stream_id < 123456789 ; [repeated for each device ID of that user, potentially a lot...] With the patch applied on my instance for the past couple of days, I no longer notice overly long statements of that particular kind. Signed-off-by: pacien <pacien.trangirard@pacien.net>	2023-07-03 16:39:38 +02:00
reivilibre	53aa26eddc	Add a timeout that aborts any Postgres statement taking more than 1 hour. (#15853 ) * Add a timeout to Postgres statements * Newsfile Signed-off-by: Olivier Wilkinson (reivilibre) <oliverw@matrix.org> --------- Signed-off-by: Olivier Wilkinson (reivilibre) <oliverw@matrix.org>	2023-07-03 11:38:57 +01:00
Shay	78cfa55dad	Fix sqlite `user_filters` upgrade (#15817 )	2023-06-27 09:41:42 +01:00
Nicolas Werner	e0c39d6bb5	Fix forgotten rooms missing in initial sync (#15815 ) If you leave a room and forget it, then rejoin it, the room would be missing from the next initial sync. fixes #13262 Signed-off-by: Nicolas Werner <n.werner@famedly.com>	2023-06-21 14:56:31 +01:00
Eric Eastwood	0f02f0b4da	Remove experimental MSC2716 implementation to incrementally import history into existing rooms (#15748 ) Context for why we're removing the implementation: - https://github.com/matrix-org/matrix-spec-proposals/pull/2716#issuecomment-1487441010 - https://github.com/matrix-org/matrix-spec-proposals/pull/2716#issuecomment-1504262734 Anyone wanting to continue MSC2716, should also address these leftover tasks: https://github.com/matrix-org/synapse/issues/10737 Closes https://github.com/matrix-org/synapse/issues/10737 in the fact that it is not longer necessary to track those things.	2023-06-16 14:12:24 -05:00
Andrew Morgan	2ac6c3bbb5	Don't always lock "user_ips" table when performing non-native upsert (#15788 )	2023-06-16 15:25:44 +01:00
Jason Little	21fea6b749	Prefill events after invalidate not before when persisting events (#15758 ) Fixes #15757	2023-06-14 09:42:18 +01:00
Shay	553f2f53e7	Replace `EventContext` fields `prev_group` and `delta_ids` with field `state_group_deltas` (#15233 )	2023-06-13 13:22:06 -07:00
Erik Johnston	c485ed1c5a	Clear event caches when we purge history (#15609 ) This should help a little with #13476 --------- Co-authored-by: Patrick Cloke <patrickc@matrix.org>	2023-06-08 13:14:40 +01:00
David Robertson	d162aecaac	Quick & dirty metric for background update status (#15740 ) * Quick & dirty metric for background update status * Changelog * Remove debug Co-authored-by: Mathieu Velten <mathieuv@matrix.org> * Actually write to _aborted --------- Co-authored-by: Mathieu Velten <mathieuv@matrix.org>	2023-06-07 17:12:23 +00:00
Eric Eastwood	e536f02f68	Remove superfluous `room_memberships` join from background update (#15733 ) Spawning from https://github.com/matrix-org/synapse/pull/15731	2023-06-07 11:47:01 -05:00
Erik Johnston	8934c11935	Merge branch 'master' into develop	2023-06-07 14:45:19 +01:00
Erik Johnston	f7c6553ebc	Fix schema delta error in 1.85 (#15739 ) Some users seem to have multiple rows per user / room with a null thread ID, which we need to handle.	2023-06-07 13:02:42 +01:00
Erik Johnston	a701c089fa	Fix schema delta error in 1.85 (#15738 ) There appears to be a race where you can end up with entries in `event_push_summary` with both a `NULL` and `main` thread ID. Fixes #15736 Introduced in #15597	2023-06-07 10:50:32 +01:00
Eric Eastwood	9d911b0da6	No need for the extra join since `membership` is built-in to `current_state_events` (#15731 ) This helps with the upstream `is_host_joined()` and `is_host_invited()` functions. `membership` was added to `current_state_events` in https://github.com/matrix-org/synapse/pull/5706 and forced in https://github.com/matrix-org/synapse/pull/13745	2023-06-06 22:19:57 -05:00
Shay	6ee96e9366	Improve performance of user directory search (#15729 )	2023-06-06 21:16:03 +01:00
Patrick Cloke	f880e64b11	Stabilize support for MSC3952: Intentional mentions. (#15520 )	2023-06-06 09:11:07 +01:00
Shay	d0c4257f14	`N + 3`: Read from column `full_user_id` rather than `user_id` of tables `profiles` and `user_filters` (#15649 )	2023-06-02 17:24:13 -07:00
Mathieu Velten	e0f2429d13	Add a catch-all * to the supported relation types when redacting (#15705 ) This is an update to MSC3912 implementation	2023-06-02 13:13:50 +00:00
H. Shay	8af29155ec	Merge branch 'release-v1.85' into develop	2023-06-01 10:26:37 -07:00
Erik Johnston	5ed0e8c61f	Cache requests for user's devices from federation (#15675 ) This should mitigate the issue where lots of different servers requests the same user's devices all at once.	2023-06-01 13:25:20 +00:00
Shay	6d9e2fd878	Speed up background jobs populate_full_user_id_user_filters and populate_full_user_id_profiles (#15700 )	2023-05-31 15:13:48 -07:00
reivilibre	11e15d79b8	Fix a performance issue introduced in Synapse v1.83.0 which meant that purging rooms was very slow and database-intensive. (#15693 ) * Add indices required to efficiently validate new foreign key constraints on stream_ordering * Newsfile Signed-off-by: Olivier Wilkinson (reivilibre) <oliverw@matrix.org> --------- Signed-off-by: Olivier Wilkinson (reivilibre) <oliverw@matrix.org>	2023-05-31 14:59:56 +01:00
Gabriel Féron	daf3a67908	Add get_canonical_room_alias to module API (#15450 ) Co-authored-by: Boxdot <d@zerovolt.org>	2023-05-31 09:18:37 -04:00
Patrick Cloke	2ad91ec628	Set thread_id column to non-null for event_push_{actions,actions_staging,summary} (#15597 ) Updates the database schema to require a thread_id (by adding a constraint that the column is non-null) for event_push_actions, event_push_actions_staging, and event_push_actions_summary. For PostgreSQL we add the constraint as NOT VALID, then VALIDATE the constraint a background job to avoid locking the table during an upgrade. Each table is updated as a separate schema delta to avoid deadlocks between them. For SQLite we simply rebuild the table & copy the data.	2023-05-26 13:16:08 -04:00
Eric Eastwood	77156a4bc1	Process previously failed backfill events in the background (#15585 ) Process previously failed backfill events in the background because they are bound to fail again and we don't need to waste time holding up the request for something that is bound to fail again. Fix https://github.com/matrix-org/synapse/issues/13623 Follow-up to https://github.com/matrix-org/synapse/issues/13621 and https://github.com/matrix-org/synapse/issues/13622 Part of making `/messages` faster: https://github.com/matrix-org/synapse/issues/13356	2023-05-24 23:22:24 -05:00
Erik Johnston	c7e9c1d5ae	Speed up user directory rebuild for users some more... (#15665 )	2023-05-24 14:13:28 +00:00
Patrick Cloke	1f55c04cbc	Improve type hints for cached decorator. (#15658 ) The cached decorators always return a Deferred, which was not properly propagated. It was close enough when wrapping coroutines, but failed if a bare function was wrapped.	2023-05-24 12:59:31 +00:00
Eric Eastwood	379eb2d7ab	Fix `@trace` not wrapping some state methods that return coroutines correctly (#15647 ) ``` 2023-05-21 09:30:09,288 - synapse.logging.opentracing - 940 - ERROR - POST-1 - @trace may not have wrapped StateStorageController.get_state_for_groups correctly! The function is not async but returned a coroutine ``` Tracing instrumentation for these functions originally introduced in https://github.com/matrix-org/synapse/pull/15610	2023-05-23 12:26:25 -05:00
Eric Eastwood	703a8f9c67	Instrument `state` and `state_group` storage related things (tracing) (#15610 ) Instrument `state` and `state_group` storage related things (tracing) so it's a little more clear where these database transactions are coming from as there is a lot of wires crossing in these functions. Part of `/messages` performance investigation: https://github.com/matrix-org/synapse/issues/13356	2023-05-19 12:26:58 -05:00
reivilibre	736199b763	Remove old R30 because R30v2 supercedes it (#10428 ) R30v2 has been out since 2021-07-19 (https://github.com/matrix-org/synapse/pull/10332) and we started collecting stats on 2021-08-16. Since it's been over a year now (almost 2 years), this is enough grace period for us to now rip it out.	2023-05-19 11:13:44 -05:00
Patrick Cloke	1e89976b26	Rename blacklist/whitelist internally. (#15620 ) Avoid renaming configuration settings for now and rename internal code to use blocklist and allowlist instead.	2023-05-19 12:25:25 +00:00
Nick Mills-Barrett	ad50510a06	Handle missing previous read marker event. (#15464 ) If the previous read marker is pointing to an event that no longer exists (e.g. due to retention) then assume that the newly given read marker is newer.	2023-05-18 14:37:31 -04:00
Patrick Cloke	375b0a8a11	Update code to refer to "workers". (#15606 ) A bunch of comments and variables are out of date and use obsolete terms.	2023-05-16 15:56:38 -04:00
Shay	9f6ff6a0eb	Add not null constraint to column `full_user_id` of tables `profiles` and `user_filters` (#15537 )	2023-05-16 10:57:39 -07:00
Eric Eastwood	c51d2e6199	Fix subscriptable type usage in Python <3.9 (#15604 ) Fix the following `mypy` errors when running `mypy` with Python 3.7: ``` synapse/storage/controllers/stats.py:58: error: "Counter" is not subscriptable, use "typing.Counter" instead [misc] tests/test_state.py:267: error: "dict" is not subscriptable, use "typing.Dict" instead [misc] ``` Part of https://github.com/matrix-org/synapse/issues/15603 In Python 3.9, `typing` is deprecated and the types are subscriptable (generics) by default, https://peps.python.org/pep-0585/#implementation	2023-05-16 12:19:46 -05:00
Erik Johnston	808105bd31	Revert "Set thread_id column to non-null for event_push_{actions,actions_staging,summary} (#15437 )" (#15580 ) This reverts commit `a7b3e9ce65`.	2023-05-12 11:38:16 +01:00
Andrew Morgan	7c95b65873	Clean up and clarify "Create or modify Account" Admin API documentation (#15544 )	2023-05-05 15:51:46 +01:00
Sean Quah	e46d5f3586	Factor out an `is_mine_server_name` method (#15542 ) Add an `is_mine_server_name` method, similar to `is_mine_id`. Ideally we would use this consistently, instead of sometimes comparing against `hs.hostname` and other times reaching into `hs.config.server.server_name`. Also fix a bug in the tests where `hs.hostname` would sometimes differ from `hs.config.server.server_name`. Signed-off-by: Sean Quah <seanq@matrix.org>	2023-05-05 15:06:22 +01:00
Erik Johnston	28ac1a1a91	Speed up deleting of old rows in `event_push_actions` (#15531 ) Enforce that we use index scans (rather than seq scans), which we also do for state queries. The reason to enforce this is that we can't correctly get PostgreSQL to understand the distribution of `stream_ordering` depends on `highlight`, and so it always defaults (on matrix.org) to sequential scans.	2023-05-03 13:42:43 +00:00
Erik Johnston	fc3a878220	Speed up rebuilding of the user directory for local users (#15529 ) The idea here is to batch up the work.	2023-05-03 13:41:37 +00:00
Patrick Cloke	a7b3e9ce65	Set thread_id column to non-null for event_push_{actions,actions_staging,summary} (#15437 ) Updates the database schema to require a thread_id (by adding a constraint that the column is non-null) for event_push_actions, event_push_actions_staging, and event_push_actions_summary. For PostgreSQL we add the constraint as NOT VALID, then VALIDATE the constraint a background job to avoid locking the table during an upgrade. For SQLite we simply rebuild the table & copy the data.	2023-05-03 07:49:03 -04:00
Sean Quah	04e79e6a18	Add config option to forget rooms automatically when users leave them (#15224 ) This is largely based off the stats and user directory updater code. Signed-off-by: Sean Quah <seanq@matrix.org>	2023-05-03 12:27:33 +01:00
Patrick Cloke	07b1c70d6b	Initial implementation of MSC3981: recursive relations API (#15315 ) Adds an optional keyword argument to the /relations API which will recurse a limited number of event relationships. This will cause the API to return not just the events related to the parent event, but also events related to those related to the parent event, etc. This is disabled by default behind an experimental configuration flag and is currently implemented using prefixed parameters.	2023-05-02 07:59:55 -04:00
Shay	89f6fb0d5a	Add an admin API endpoint to support per-user feature flags (#15344 )	2023-04-28 11:33:45 -07:00
Patrick Cloke	57aeeb308b	Add support for claiming multiple OTKs at once. (#15468 ) MSC3983 provides a way to request multiple OTKs at once from appservices, this extends this concept to the Client-Server API. Note that this will likely be spit out into a separate MSC, but is currently part of MSC3983.	2023-04-27 12:57:46 -04:00
Patrick Cloke	6efa674004	Add type hints to schema deltas (#15497 ) Cleans-up the schema delta files: * Removes no-op functions. * Adds missing type hints to function parameters. * Fixes any issues with type hints. This also renames one (very old) schema delta to avoid a conflict that mypy complains about.	2023-04-27 12:44:53 +00:00
Patrick Cloke	a346b43837	Check databases/__init__ and main/cache with mypy. (#15496 )	2023-04-27 07:59:14 -04:00
Shay	301b4156d5	Add column `full_user_id` to tables `profiles` and `user_filters`. (#15458 )	2023-04-26 16:03:26 -07:00
Erik Johnston	9900f7c231	Add admin endpoint to query room sizes (#15482 )	2023-04-26 16:00:11 +00:00
Patrick Cloke	8e9739449d	Add unstable /keys/claim endpoint which always returns fallback keys. (#15462 ) It can be useful to always return the fallback key when attempting to claim keys. This adds an unstable endpoint for `/keys/claim` which always returns fallback keys in addition to one-time-keys. The fallback key(s) are not marked as "used" unless there are no corresponding OTKs. This is currently defined in MSC3983 (although likely to be split out to a separate MSC). The endpoint shape may change or be requested differently (i.e. a keyword parameter on the current endpoint), but the core logic should be reasonable.	2023-04-25 13:30:41 -04:00
Nick Mills-Barrett	c55293c230	Re re introduce membership tables event stream ordering (#15356 )	2023-04-25 09:44:29 +01:00
Quentin Gliech	8b3a502996	Experimental support for MSC3970: per-device transaction IDs (#15318 )	2023-04-25 09:37:09 +01:00
Patrick Cloke	5e024a0645	Modify StoreKeyFetcher to read from server_keys_json. (#15417 ) Before this change: * `PerspectivesKeyFetcher` and `ServerKeyFetcher` write to `server_keys_json`. * `PerspectivesKeyFetcher` also writes to `server_signature_keys`. * `StoreKeyFetcher` reads from `server_signature_keys`. After this change: * `PerspectivesKeyFetcher` and `ServerKeyFetcher` write to `server_keys_json`. * `PerspectivesKeyFetcher` also writes to `server_signature_keys`. * `StoreKeyFetcher` reads from `server_keys_json`. This results in `StoreKeyFetcher` now using the results from `ServerKeyFetcher` in addition to those from `PerspectivesKeyFetcher`, i.e. keys which are directly fetched from a server will now be pulled from the database instead of refetched. An additional minor change is included to avoid creating a `PerspectivesKeyFetcher` (and checking it) if no `trusted_key_servers` are configured. The overall impact of this should be better usage of cached results: * If a server has no trusted key servers configured then it should reduce how often keys are fetched. * if a server's trusted key server does not have a requested server's keys cached then it should reduce how often keys are directly fetched.	2023-04-20 12:30:32 -04:00
David Robertson	8a47d6e3a6	More precise type for LoggingTransaction.execute (#15432 ) * More precise type for LoggingTransaction.execute * Add an annotation for stream_ordering_month_ago This would have spotted the error that was fixed in "Add comma missing from #15382. (#15429)"	2023-04-14 18:04:49 +00:00
Erik Johnston	b5192355f6	User directory background update speedup (#15435 ) c.f. #15264 The two changes are: 1. Add indexes so that the select / deletes don't do sequential scans 2. Don't repeatedly call `SELECT count(*)` each iteration, as that's slow	2023-04-14 16:10:32 +01:00
Dirk Klimpel	4af0aec54d	Load `/directory/room/{roomAlias}` endpoint on workers (#15333 ) * Enable `directory` * move to worker store * newsfile * disable `ClientDirectoryListServer` and `ClientAppserviceDirectoryListServer` for workers	2023-04-14 10:24:06 +01:00
reivilibre	edae20f926	Improve robustness when handling a perspective key response by deduplicating received server keys. (#15423 ) * Change `store_server_verify_keys` to take a `Mapping[(str, str), FKR]` This is because we already can't handle duplicate keys — leads to cardinality violation * Newsfile Signed-off-by: Olivier Wilkinson (reivilibre) <oliverw@matrix.org> --------- Signed-off-by: Olivier Wilkinson (reivilibre) <oliverw@matrix.org>	2023-04-13 15:35:03 +01:00
reivilibre	38272be037	Add comma missing from #15382 . (#15429 ) * Add missing comma * Newsfile Signed-off-by: Olivier Wilkinson (reivilibre) <oliverw@matrix.org> --------- Signed-off-by: Olivier Wilkinson (reivilibre) <oliverw@matrix.org>	2023-04-13 15:06:25 +01:00
Patrick Cloke	d07d255830	Implement MSC2175: remove the creator field from create events. (#15394 )	2023-04-06 16:26:28 -04:00
Erik Johnston	485b9fdefb	Don't keep old stream_ordering_to_exterm around (#15382 )	2023-04-06 16:42:39 +00:00
Patrick Cloke	72b43bec8b	Merge remote-tracking branch 'origin/release-v1.81' into develop	2023-04-06 11:44:26 -04:00
Quentin Gliech	6eb3edec47	Fix the 'set_device_id_for_pushers_txn' background update. (#15391 ) Refer to the correct field from the response when updating the background update progress.	2023-04-05 07:49:15 -04:00
Shay	6b23d74ad1	Delete server-side backup keys when deactivating an account. (#15181 )	2023-04-04 20:16:08 +00:00
Erik Johnston	79d2e2e79c	Speed up membership queries for users with forgotten rooms (#15385 )	2023-04-04 14:11:34 +01:00
Erik Johnston	6204c3663e	Revert pruning of old devices (#15360 ) * Revert "Fix registering a device on an account with lots of devices (#15348)" This reverts commit `f0d8f66eaa`. * Revert "Delete stale non-e2e devices for users, take 3 (#15183)" This reverts commit `78cdb72cd6`.	2023-03-31 13:51:51 +01:00
Olivier Wilkinson (reivilibre)	72d2ceaa9a	Revert "Set thread_id column to non-null for event_push_{actions,actions_staging,summary} (#15350 )" This reverts commit `2a234b788e`. See #15359 for context.	2023-03-31 12:10:10 +01:00
Patrick Cloke	2a234b788e	Set thread_id column to non-null for event_push_{actions,actions_staging,summary} (#15350 ) Clean-up from adding the thread_id column, which was initially null but backfilled with values. It is desirable to require it to now be non-null. In addition to altering this column to be non-null, we clean up obsolete background jobs, indexes, and just-in-time updating code.	2023-03-30 15:11:31 -04:00
Mathieu Velten	6f68e32bfb	to_device updates could be dropped when consuming the replication stream (#15349 ) Co-authored-by: reivilibre <oliverw@matrix.org>	2023-03-30 19:41:14 +02:00
Erik Johnston	91c3f32673	Speed up SQLite unit test CI (#15334 ) Tests now take 40% of the time.	2023-03-30 16:21:12 +01:00
Sean Quah	d9f694932c	Fix spinloop during partial state sync when a prev event is in backoff (#15351 ) Previously, we would spin in a tight loop until `update_state_for_partial_state_event` stopped raising `FederationPullAttemptBackoffError`s. Replace the spinloop with a wait until the backoff period has expired. Signed-off-by: Sean Quah <seanq@matrix.org>	2023-03-30 13:36:41 +01:00
Erik Johnston	f0d8f66eaa	Fix registering a device on an account with lots of devices (#15348 ) Fixes up #15183	2023-03-29 13:37:06 +00:00
Erik Johnston	5350b5d04d	Revert "Reintroduce membership tables event stream ordering (#15128 )" (#15347 ) This reverts commit `e6af49fbea`.	2023-03-29 13:24:28 +01:00
Erik Johnston	78cdb72cd6	Delete stale non-e2e devices for users, take 3 (#15183 ) This should help reduce the number of devices e.g. simple bots the repeatedly login rack up. We only delete non-e2e devices as they should be safe to delete, whereas if we delete e2e devices for a user we may accidentally break their ability to receive e2e keys for a message.	2023-03-29 12:07:14 +01:00
Patrick Cloke	5282ba1e2b	Implement MSC3983 to proxy /keys/claim queries to appservices. (#15314 ) Experimental support for MSC3983 is behind a configuration flag. If enabled, for users which are exclusively owned by an application service then the appservice will be queried for one-time keys if there are none uploaded to Synapse.	2023-03-28 18:26:27 +00:00
dependabot[bot]	bd4d958aaf	Bump ruff from 0.0.252 to 0.0.259 (#15328 ) * Bump ruff from 0.0.252 to 0.0.259 Bumps [ruff](https://github.com/charliermarsh/ruff) from 0.0.252 to 0.0.259. - [Release notes](https://github.com/charliermarsh/ruff/releases) - [Changelog](https://github.com/charliermarsh/ruff/blob/main/BREAKING_CHANGES.md) - [Commits](https://github.com/charliermarsh/ruff/compare/v0.0.252...v0.0.259) --- updated-dependencies: - dependency-name: ruff dependency-type: direct:development update-type: version-update:semver-patch ... Signed-off-by: dependabot[bot] <support@github.com> * Fix new warnings * Mypy * Newsfile --------- Signed-off-by: dependabot[bot] <support@github.com> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com> Co-authored-by: Erik Johnston <erik@matrix.org>	2023-03-28 09:46:47 +01:00
reivilibre	5f7c908280	As an optimisation, use `TRUNCATE` on Postgres when clearing the user directory tables. (#15316 )	2023-03-24 15:31:12 +00:00
Quentin Gliech	5b70f240cf	Make cleaning up pushers depend on the device_id instead of the token_id (#15280 ) This makes it so that we rely on the `device_id` to delete pushers on logout, instead of relying on the `access_token_id`. This ensures we're not removing pushers on token refresh, and prepares for a world without access token IDs (also known as the OIDC). This actually runs the `set_device_id_for_pushers` background update, which was forgotten in #13831. Note that for backwards compatibility it still deletes pushers based on the `access_token` until the background update finishes.	2023-03-24 11:09:39 -04:00
Nick Mills-Barrett	e6af49fbea	Reintroduce membership tables event stream ordering (#15128 ) * Add `event_stream_ordering` column to membership state tables Specifically this adds the column to `current_state_events`, `local_current_membership` and `room_memberships`. Each of these tables is regularly joined with the `events` table to get the stream ordering and denormalising this into each table will yield significant query performance improvements once used. * Make denormalised `event_stream_ordering` columns foreign keys * Add comment in schema file explaining new denormalised columns * Add triggers to enforce consistency of `event_stream_ordering` columns * Re-order purge room tables to account for foreign keys * Bump schema version to 75 Co-authored-by: David Robertson <david.m.robertson1@gmail.com> Co-authored-by: Richard van der Hoff <1389908+richvdh@users.noreply.github.com>	2023-03-24 11:44:01 +00:00
David Robertson	3b0083c92a	Use immutabledict instead of frozendict (#15113 ) Additionally: * Consistently use `freeze()` in test --------- Co-authored-by: Patrick Cloke <clokep@users.noreply.github.com> Co-authored-by: 6543 <6543@obermui.de>	2023-03-22 17:15:34 +00:00
Patrick Cloke	1bc4feb6c9	Apply & bundle edits for non-message events. (#15295 )	2023-03-21 14:19:54 -04:00
Andrew Morgan	ec9224bf9a	Make `POST /_matrix/client/v3/rooms/{roomId}/report/{eventId}` endpoint return 404 if event exists, but the user lacks access (#15300 )	2023-03-21 13:24:03 +00:00
reivilibre	1f5473465d	Refresh remote profiles that have been marked as stale, in order to fill the user directory. [rei:userdirpriv] (#14756 ) * Scaffolding for background process to refresh profiles * Add scaffolding for background process to refresh profiles for a given server * Implement the code to select servers to refresh from * Ensure we don't build up multiple looping calls * Make `get_profile` able to respect backoffs * Add logic for refreshing users * When backing off, schedule a refresh when the backoff is over * Wake up the background processes when we receive an interesting state event * Add tests * Newsfile Signed-off-by: Olivier Wilkinson (reivilibre) <oliverw@matrix.org> * Add comment about 1<<62 --------- Signed-off-by: Olivier Wilkinson (reivilibre) <oliverw@matrix.org>	2023-03-16 11:44:11 +00:00
reivilibre	f54f877f27	Preparatory work to fix the user directory assuming that any remote membership state events represent a profile change. [rei:userdirpriv] (#14755 ) * Remove special-case method for new memberships only, use more generic method * Only collect profiles from state events in public rooms * Add a table to track stale remote user profiles * Add store methods to set and delete rows in this new table * Mark remote profiles as stale when a member state event comes in to a private room * Newsfile Signed-off-by: Olivier Wilkinson (reivilibre) <oliverw@matrix.org> * Simplify by removing Optionality of `event_id` * Replace names and avatars with None if they're set to dodgy things I think this makes more sense anyway. * Move schema delta to 74 (I missed the boat?) * Turns out these can be None after all --------- Signed-off-by: Olivier Wilkinson (reivilibre) <oliverw@matrix.org>	2023-03-16 09:55:19 +00:00
reivilibre	63d87c08c8	Add schema comments about the `destinations` and `destination_rooms` tables. (#15247 )	2023-03-15 09:25:58 +00:00
reivilibre	d0fe417f5c	Remove unused store method `_set_destination_retry_timings_emulated`. (#15266 )	2023-03-14 17:32:46 +00:00
Patrick Cloke	3d060eae6c	Add missing type hints to `synapse.storage.database`. (#15230 )	2023-03-09 07:10:09 -05:00
Patrick Cloke	88efc75bab	Include the room ID in more purge room log lines. (#15222 )	2023-03-08 20:08:56 +00:00
Erik Johnston	c69aae94cd	Split up txn for fetching device keys (#15215 ) We look up keys in batches, but we should do that outside of the transaction to avoid starving the database pool.	2023-03-07 08:51:34 +00:00
Patrick Cloke	02f74f3a99	Combine AbstractStreamIdTracker and AbstractStreamIdGenerator. (#15192 ) AbstractStreamIdTracker (now) has only a single sub-class: AbstractStreamIdGenerator, combine them to simplify some code and remove any direct references to AbstractStreamIdTracker.	2023-03-03 08:13:37 -05:00
Andrew Morgan	15e975f68f	Experimental MSC3890 Implementation: Fix deleting account data when using an account data writer worker (#14869 )	2023-03-03 10:51:57 +00:00
Andrew Morgan	1eea662780	Add a `get_next_txn` method to `StreamIdGenerator` to match `MultiWriterIdGenerator` (#15191	2023-03-02 18:27:00 +00:00
Dirk Klimpel	65f10afb64	Move event_reports to `RoomWorkerStore` (#15165 )	2023-03-02 10:38:46 +00:00
Richard van der Hoff	2b78981736	Remove support for aggregating reactions (#15172 ) It turns out that no clients rely on server-side aggregation of `m.annotation` relationships: it's just not very useful as currently implemented. It's also non-trivial to calculate. I want to remove it from MSC2677, so to keep the implementation in line, let's remove it here.	2023-02-28 18:49:28 +00:00
reivilibre	d62cd940cb	Fix a long-standing bug where an initial sync would not respond to changes to the list of ignored users if there was an initial sync cached. (#15163 )	2023-02-28 17:11:26 +00:00
reivilibre	682d31c702	Allow use of the `/filter` Client-Server APIs on workers. (#15134 )	2023-02-28 16:37:19 +00:00
Dirk Klimpel	93f7955eba	Admin API endpoint to delete a reported event (#15116 ) * Admin api to delete event report * lint + tests * newsfile * Apply suggestions from code review Co-authored-by: David Robertson <david.m.robertson1@gmail.com> * revert changes - move to WorkerStore * update unit test * Note that timestamp is in millseconds --------- Co-authored-by: David Robertson <david.m.robertson1@gmail.com>	2023-02-28 12:09:10 +00:00
Andrew Morgan	b40657314e	Add module API callbacks for adding and deleting local 3PID associations (#15044	2023-02-27 14:19:19 +00:00
Shay	1c95ddd09b	Batch up storing state groups when creating new room (#14918 )	2023-02-24 13:15:29 -08:00
Sean Quah	335f52d595	Improve handling of non-ASCII characters in user directory search (#15143 ) * Fix a long-standing bug where non-ASCII characters in search terms, including accented letters, would not match characters in a different case. * Fix a long-standing bug where search terms using combining accents would not match display names using precomposed accents and vice versa. To fully take effect, the user directory must be rebuilt after this change. Fixes #14630. Signed-off-by: Sean Quah <seanq@matrix.org>	2023-02-24 13:39:45 +00:00
dependabot[bot]	9bb2eac719	Bump black from 22.12.0 to 23.1.0 (#15103 )	2023-02-22 15:29:09 -05:00
reivilibre	1cbc3f197c	Fix a bug introduced in Synapse v1.74.0 where searching with colons when using ICU for search term tokenisation would fail with an error. (#15079 ) Co-authored-by: David Robertson <davidr@element.io>	2023-02-20 12:00:18 +00:00
David Robertson	ffc2ee521d	Use mypy 1.0 (#15052 ) * Update mypy and mypy-zope * Remove unused ignores These used to suppress ``` synapse/storage/engines/__init__.py:28: error: "__new__" must return a class instance (got "NoReturn") [misc] ``` and ``` synapse/http/matrixfederationclient.py:1270: error: "BaseException" has no attribute "reasons" [attr-defined] ``` (note that we check `hasattr(e, "reasons")` above) * Avoid empty body warnings, sometimes by marking methods as abstract E.g. ``` tests/handlers/test_register.py:58: error: Missing return statement [empty-body] tests/handlers/test_register.py:108: error: Missing return statement [empty-body] ``` * Suppress false positive about `JaegerConfig` Complaint was ``` synapse/logging/opentracing.py:450: error: Function "Type[Config]" could always be true in boolean context [truthy-function] ``` * Fix not calling `is_state()` Oops! ``` tests/rest/client/test_third_party_rules.py:428: error: Function "Callable[[], bool]" could always be true in boolean context [truthy-function] ``` * Suppress false positives from ParamSpecs ```` synapse/logging/opentracing.py:971: error: Argument 2 to "_custom_sync_async_decorator" has incompatible type "Callable[[Arg(Callable[P, R], 'func'), P], _GeneratorContextManager[None]]"; expected "Callable[[Callable[P, R], P], _GeneratorContextManager[None]]" [arg-type] synapse/logging/opentracing.py:1017: error: Argument 2 to "_custom_sync_async_decorator" has incompatible type "Callable[[Arg(Callable[P, R], 'func'), P], _GeneratorContextManager[None]]"; expected "Callable[[Callable[P, R], P], _GeneratorContextManager[None]]" [arg-type] ```` * Drive-by improvement to `wrapping_logic` annotation * Workaround false "unreachable" positives See https://github.com/Shoobx/mypy-zope/issues/91 ``` tests/http/test_proxyagent.py:626: error: Statement is unreachable [unreachable] tests/http/test_proxyagent.py:762: error: Statement is unreachable [unreachable] tests/http/test_proxyagent.py:826: error: Statement is unreachable [unreachable] tests/http/test_proxyagent.py:838: error: Statement is unreachable [unreachable] tests/http/test_proxyagent.py:845: error: Statement is unreachable [unreachable] tests/http/federation/test_matrix_federation_agent.py:151: error: Statement is unreachable [unreachable] tests/http/federation/test_matrix_federation_agent.py:452: error: Statement is unreachable [unreachable] tests/logging/test_remote_handler.py:60: error: Statement is unreachable [unreachable] tests/logging/test_remote_handler.py:93: error: Statement is unreachable [unreachable] tests/logging/test_remote_handler.py:127: error: Statement is unreachable [unreachable] tests/logging/test_remote_handler.py:152: error: Statement is unreachable [unreachable] ``` * Changelog * Tweak DBAPI2 Protocol to be accepted by mypy 1.0 Some extra context in: - https://github.com/matrix-org/python-canonicaljson/pull/57 - https://github.com/python/mypy/issues/6002 - https://mypy.readthedocs.io/en/latest/common_issues.html#covariant-subtyping-of-mutable-protocol-members-is-rejected * Pull in updated canonicaljson lib so the protocol check just works * Improve comments in opentracing I tried to workaround the ignores but found it too much trouble. I think the corresponding issue is https://github.com/python/mypy/issues/12909. The mypy repo has a PR claiming to fix this (https://github.com/python/mypy/pull/14677) which might mean this gets resolved soon? * Better annotation for INTERACTIVE_AUTH_CHECKERS * Drive-by AUTH_TYPE annotation, to remove an ignore	2023-02-16 16:09:11 +00:00
David Robertson	06ba71083e	Fix order of partial state tables when purging (#15068 ) * Fix order of partial state tables when purging `partial_state_rooms` has an FK on `events` pointing to the join event we get from `/send_join`, so we must delete from that table before deleting from `events`. NB: It would be nice to cancel any resync processes for the room being purged. We do not do this at present. To do so reliably we'd need an internal HTTP "replication" endpoint, because the worker doing the resync process may be different to that handling the purge request. The first time the resync process tries to write data after the deletion it will fail because we have deleted necessary data e.g. auth events. AFAICS it will not retry the resync, so the only downside to not cancelling the resync is a scary-looking traceback. (This is presumably extremely race-sensitive.) * Changelog * admist(?) -> between * Warn about a race * Fix typo, thanks Sean Co-authored-by: Sean Quah <8349537+squahtx@users.noreply.github.com> --------- Co-authored-by: Sean Quah <8349537+squahtx@users.noreply.github.com>	2023-02-14 23:42:29 +00:00
Erik Johnston	cb262713b7	Fix clashing DB txn name (#15070 ) * Fix clashing DB txn name * Newsfile	2023-02-14 11:20:25 +00:00
Erik Johnston	f09db5c991	Skip calculating unread push actions in `/sync` when `enable_push` is false. (#14980 )	2023-02-14 11:10:29 +00:00
Harishankar Kumar	db2b105d69	Change collection[str] to StrCollection in event_auth code (#14929 ) Signed-off-by: Harishankar Kumar <hari01584@gmail.com>	2023-02-14 09:37:08 +00:00
Mathieu Velten	6cddf24e36	Faster joins: don't stall when a user joins during a fast join (#14606 ) Fixes #12801. Complement tests are at https://github.com/matrix-org/complement/pull/567. Avoid blocking on full state when handling a subsequent join into a partial state room. Also always perform a remote join into partial state rooms, since we do not know whether the joining user has been banned and want to avoid leaking history to banned users. Signed-off-by: Mathieu Velten <mathieuv@matrix.org> Co-authored-by: Sean Quah <seanq@matrix.org> Co-authored-by: David Robertson <davidr@element.io>	2023-02-10 23:31:05 +00:00
Sean Quah	d0c713cc85	Return read-only collections from `@cached` methods (#13755 ) It's important that collections returned from `@cached` methods are not modified, otherwise future retrievals from the cache will return the modified collection. This applies to the return values from `@cached` methods and the values inside the dictionaries returned by `@cachedList` methods. It's not necessary for the dictionaries returned by `@cachedList` methods themselves to be read-only. Signed-off-by: Sean Quah <seanq@matrix.org> Co-authored-by: David Robertson <davidr@element.io>	2023-02-10 23:29:00 +00:00
Patrick Cloke	cf5233b783	Avoid fetching unused account data in sync. (#14973 ) The per-room account data is no longer unconditionally fetched, even if all rooms will be filtered out. Global account data will not be fetched if it will all be filtered out.	2023-02-10 14:22:16 +00:00
David Robertson	d793fcd241	Merge branch 'release-v1.77' into develop	2023-02-10 13:43:18 +00:00
Patrick Cloke	a481fb9f98	Refactor get_user_devices_from_cache to avoid mutating cached values. (#15040 ) The previous version of the code could mutate a cached value, but only if the input requested all devices of a user and a specific device. To avoid this nonsensical situation we no longer fetch a specific device ID if all of a user's devices are returned.	2023-02-10 08:09:47 -05:00
Erik Johnston	fd296b7343	Fix exception on start up about device lists (#15041 ) Fixes #15010.	2023-02-10 09:52:35 +00:00
Andrew Morgan	c1d2ce2901	Do not always start a db txn on Postgres (#14840 )	2023-02-09 19:57:01 +00:00
David Robertson	cd2484dc2e	Bump schema version (#15036 ) * Bump schema version This should have been included in `f10caa73ee` (and #14979). * Changelog	2023-02-09 15:28:26 +00:00
Patrick Cloke	733531ee3e	Add final type hint to synapse.server. (#15035 )	2023-02-09 09:49:04 -05:00
David Robertson	f10caa73ee	Disambiguate `get_ex_outlier_stream_rows` query A backwards-compatible piece of #14979 that's safe to land now.	2023-02-07 15:33:33 +00:00
David Robertson	9cd7610f86	Revert "Add `event_stream_ordering` column to membership state tables (#14979 )" This reverts commit `5fdc12f482`.	2023-02-07 15:26:55 +00:00
Nick Mills-Barrett	5fdc12f482	Add `event_stream_ordering` column to membership state tables (#14979 ) This adds an `event_stream_ordering` column to `current_state_events`, `local_current_membership` and `room_memberships`. Each of these tables is regularly joined with the `events` table to get the stream ordering and denormalising this into each table will yield significant query performance improvements once used. Includes a background job to populate these values from the `events` table. Same idea as https://github.com/matrix-org/synapse/pull/13703. Signed off by Nick @ Beeper (@fizzadar).	2023-02-07 00:10:54 +00:00
David Robertson	e8269ed391	Type hints for tests.appservice (#14990 ) * Accept a Sequence of events in synapse.appservice This avoids some casts/ignores in the tests I'm about to fixup. It seems that `List[Mock]` is not a subtype of `List[EventBase]`, but `Sequence[Mock]` is a subtype of `Sequence[EventBase]`. So presumably `Mock` is considered a subtype of anything, much like `Any`. * make tests.appservice.test_scheduler pass mypy * Extra hints in tests.appservice.test_scheduler * Extra hints in tests.appservice.test_api * Extra hints in tests.appservice.test_appservice * Disallow untyped defs * Changelog	2023-02-06 12:49:06 +00:00
Patrick Cloke	b2d97bac09	Implement MSC3958: suppress notifications from edits (#14960 ) Co-authored-by: Brad Murray <brad@beeper.com> Co-authored-by: Nick Barrett <nick@beeper.com> Copy the suppress_edits push rule from Beeper to implement MSC3958. `9415a1284b/rust/src/push/base_rules.rs (L98-L114)`	2023-02-03 14:31:14 -05:00
Sean Quah	0a686d1d13	Faster joins: Refactor handling of servers in room (#14954 ) Ensure that the list of servers in a partial state room always contains the server we joined off. Also refactor `get_partial_state_servers_at_join` to return `None` when the given room is no longer partial stated, to explicitly indicate when the room has partial state. Otherwise it's not clear whether an empty list means that the room has full state, or the room is partial stated, but the server we joined off told us that there are no servers in the room. Signed-off-by: Sean Quah <seanq@matrix.org>	2023-02-03 15:39:59 +00:00
David Robertson	2186ebed6c	Fetch fewer events when getting hosts in room (#14962 )	2023-02-02 16:49:14 +00:00
Patrick Cloke	1182ae5063	Add helper to parse an enum from query args & use it. (#14956 ) The `parse_enum` helper pulls an enum value from the query string (by delegating down to the parse_string helper with values generated from the enum). This is used to pull out "f" and "b" in most places and then we thread the resulting Direction enum throughout more code.	2023-02-01 21:35:24 +00:00
Patrick Cloke	230a831c73	Attempt to delete more duplicate rows in receipts_linearized table. (#14915 ) The previous assumption was that the stream_id column was unique (for a room ID, receipt type, user ID tuple), but this turned out to be incorrect. Now find the max stream ID, then map this back to a database-specific row identifier and delete other rows which match the (room ID, receipt type, user ID) tuple, but not the row ID.	2023-02-01 15:45:10 -05:00
Sean Quah	6d14fdc271	Make sqlite database migrations transactional again, part two (#14926 ) #14910 fixed the regression introduced by #13873 where sqlite database migrations would no longer run inside a transaction. However, it committed the transaction before Synapse updated its bookkeeping of which migrations have been run, which means that migrations may be run again after they have completed successfully. Leave the transaction open at the end of `executescript`, to restore the old, correct behaviour. Also make the PostgreSQL behaviour consistent with SQLite. Fixes #14909. Signed-off-by: Sean Quah <seanq@matrix.org>	2023-01-31 11:03:55 +00:00
David Robertson	796a4b7482	Prefer `type(x) is int` to `isinstance(x, int)` (#14945 ) * Perfer `type(x) is int` to `isinstance(x, int)` This covered all additional instances I could see where `x` was user-controlled. The remaining cases are ``` $ rg -s 'isinstance.[^_]int' tests/replication/_base.py 576: if isinstance(obj, int): synapse/util/caches/stream_change_cache.py 136: assert isinstance(stream_pos, int) 214: assert isinstance(stream_pos, int) 246: assert isinstance(stream_pos, int) 267: assert isinstance(stream_pos, int) synapse/replication/tcp/external_cache.py 133: if isinstance(result, int): synapse/metrics/__init__.py 100: if isinstance(calls, (int, float)): synapse/handlers/appservice.py 262: assert isinstance(new_token, int) synapse/config/_util.py 62: if isinstance(p, int): ``` which cover metrics, logic related to `jsonschema`, and replication and data streams. AFAICS these are all internal to Synapse Changelog	2023-01-31 10:33:07 +00:00
Patrick Cloke	2a51f3ec36	Implement MSC3952: Intentional mentions (#14823 ) MSC3952 defines push rules which searches for mentions in a list of Matrix IDs in the event body, instead of searching the entire event body for display name / local part. This is implemented behind an experimental configuration flag and does not yet implement the backwards compatibility pieces of the MSC.	2023-01-27 10:16:21 -05:00
David Robertson	faecc6c083	Merge branch 'release-v1.76' into develop	2023-01-27 13:01:18 +00:00
Patrick Cloke	265735db9d	Use an enum for direction. (#14927 ) For better type safety we use an enum instead of strings to configure direction (backwards or forwards).	2023-01-27 07:27:55 -05:00
Patrick Cloke	345576bc34	Fix paginating /relations with a live token (#14866 ) The `/relations` endpoint was not properly handle "live tokens" (i.e sync tokens), to do this properly we abstract the code that `/messages` has and re-use it.	2023-01-26 13:24:15 -05:00
Patrick Cloke	8a05d5de21	Batch look-ups to see if rooms are partial stated. (#14917 ) * Batch look-ups to see if rooms are partial stated. * Fix issues found in linting. * Fix typo. * Apply suggestions from code review Co-authored-by: Sean Quah <8349537+squahtx@users.noreply.github.com> * Clarify comments. Co-authored-by: Sean Quah <8349537+squahtx@users.noreply.github.com> * Also improve the cache size while we're at it * is_partial_state_rooms -> is_partial_state_room_batched * Run `black` * Improve annotation for `simple_select_many_batch` * Fix is_partial_state_room_batched impl * Okay, _actually_ fix impl * Update description. * Update synapse/storage/databases/main/room.py Co-authored-by: Patrick Cloke <clokep@users.noreply.github.com> * Run black. Co-authored-by: Sean Quah <8349537+squahtx@users.noreply.github.com> Co-authored-by: David Robertson <davidr@element.io>	2023-01-26 17:15:36 +00:00
Sean Quah	cf66d712c6	Fix initialization of `_device_list_id_gen` (#14914 ) On startup, the `_device_list_id_gen` stream id generator is initialized using the maximum stream id seen in a list of tables. When we started populating the `device_list_remote_pending` table in #13913, we forgot to add it to the aforementioned list of tables, so the stream id generator can hand out old stream ids after a restart. The end result is that Synapse can fail to handle device list update EDUs after a restart when a partial state join is in progress. Add the `device_list_remote_pending` table to the list of tables to consider when initializing the `_device_list_id_gen` stream id generator. Signed-off-by: Sean Quah <seanq@matrix.org>	2023-01-26 10:38:49 +00:00
Sean Quah	a63d4cc9e9	Make sqlite database migrations transactional again (#14910 ) #13873 introduced a regression which causes sqlite database migrations to no longer run inside a transaction. Wrap them in a transaction again, to avoid database corruption when migrations are interrupted. Fixes #14909. Signed-off-by: Sean Quah <seanq@matrix.org>	2023-01-25 13:38:53 +00:00
David Robertson	4607be0b7b	Request partial joins by default (#14905 ) * Request partial joins by default This is a little sloppy, but we are trying to gain confidence in faster joins in the upcoming RC. Admins can still opt out by adding the following to their Synapse config: ```yaml experimental: faster_joins: false ``` We may revert this change before the release proper, depending on how testing in the wild goes. * Changelog * Try to fix the backfill test failures * Upgrade notes * Postgres compat?	2023-01-24 15:28:20 +00:00
David Robertson	80d44060c9	Faster joins: omit partial rooms from eager syncs until the resync completes (#14870 ) * Allow `AbstractSet` in `StrCollection` Or else frozensets are excluded. This will be useful in an upcoming commit where I plan to change a function that accepts `List[str]` to accept `StrCollection` instead. * `rooms_to_exclude` -> `rooms_to_exclude_globally` I am about to make use of this exclusion mechanism to exclude rooms for a specific user and a specific sync. This rename helps to clarify the distinction between the global config and the rooms to exclude for a specific sync. * Better function names for internal sync methods * Track a list of excluded rooms on SyncResultBuilder I plan to feed a list of partially stated rooms for this sync to ignore * Exclude partial state rooms during eager sync using the mechanism established in the previous commit * Track un-partial-state stream in sync tokens So that we can work out which rooms have become fully-stated during a given sync period. * Fix mutation of `@cached` return value This was fouling up a complement test added alongside this PR. Excluding a room would mean the set of forgotten rooms in the cache would be extended. This means that room could be erroneously considered forgotten in the future. Introduced in #12310, Synapse 1.57.0. I don't think this had any user-visible side effects (until now). * SyncResultBuilder: track rooms to force as newly joined Similar plan as before. We've omitted rooms from certain sync responses; now we establish the mechanism to reintroduce them into future syncs. * Read new field, to present rooms as newly joined * Force un-partial-stated rooms to be newly-joined for eager incremental syncs only, provided they're still fully stated * Notify user stream listeners to wake up long polling syncs * Changelog * Typo fix Co-authored-by: Sean Quah <8349537+squahtx@users.noreply.github.com> * Unnecessary list cast Co-authored-by: Sean Quah <8349537+squahtx@users.noreply.github.com> * Rephrase comment Co-authored-by: Sean Quah <8349537+squahtx@users.noreply.github.com> * Another comment Co-authored-by: Sean Quah <8349537+squahtx@users.noreply.github.com> * Fixup merge(?) * Poke notifier when receiving un-partial-stated msg over replication * Fixup merge whoops Thanks MV :) Co-authored-by: Mathieu Velen <mathieuv@matrix.org> Co-authored-by: Mathieu Velten <mathieuv@matrix.org> Co-authored-by: Sean Quah <8349537+squahtx@users.noreply.github.com>	2023-01-23 15:44:39 +00:00
Patrick Cloke	82d3efa312	Skip processing stats for broken rooms. (#14873 ) * Skip processing stats for broken rooms. * Newsfragment * Use a custom exception.	2023-01-23 11:36:20 +00:00
Sean Quah	2ec9c58496	Faster joins: Update room stats and the user directory on workers when finishing join (#14874 ) * Faster joins: Update room stats and user directory on workers when done When finishing a partial state join to a room, we update the current state of the room without persisting additional events. Workers receive notice of the current state update over replication, but neglect to wake the room stats and user directory updaters, which then get incidentally triggered the next time an event is persisted or an unrelated event persister sends out a stream position update. We wake the room stats and user directory updaters at the appropriate time in this commit. Part of #12814 and #12815. Signed-off-by: Sean Quah <seanq@matrix.org> * fixup comment Signed-off-by: Sean Quah <seanq@matrix.org>	2023-01-23 10:31:36 +00:00
reivilibre	22cc93afe3	Enable Faster Remote Room Joins against worker-mode Synapse. (#14752 ) * Enable Complement tests for Faster Remote Room Joins on worker-mode * (dangerous) Add an override to allow Complement to use FRRJ under workers * Newsfile Signed-off-by: Olivier Wilkinson (reivilibre) <oliverw@matrix.org> * Fix race where we didn't send out replication notification * MORE HACKS * Fix get_un_partial_stated_rooms_token to take instance_name * Fix bad merge * Remove warning * Correctly advance un_partial_stated_room_stream * Fix merge * Add another notify_replication * Fixups * Create a separate ReplicationNotifier * Fix test * Fix portdb * Create a separate ReplicationNotifier * Fix test * Fix portdb * Fix presence test * Newsfile * Apply suggestions from code review * Update changelog.d/14752.misc Co-authored-by: Erik Johnston <erik@matrix.org> * lint Signed-off-by: Olivier Wilkinson (reivilibre) <oliverw@matrix.org> Co-authored-by: Erik Johnston <erik@matrix.org>	2023-01-22 21:10:11 +00:00
Erik Johnston	65d0386693	Always notify replication when a stream advances (#14877 ) This ensures that all other workers are told about stream updates in a timely manner, without having to remember to manually poke replication.	2023-01-20 18:02:18 +00:00
Andrew Morgan	a7b54ca8d8	Implement MSC3930: polls push rules (#14787 )	2023-01-19 12:47:10 +00:00
Erik Johnston	9187fd940e	Wait for streams to catch up when processing HTTP replication. (#14820 ) This should hopefully mitigate a class of races where data gets out of sync due a HTTP replication request racing with the replication streams.	2023-01-18 19:35:29 +00:00
Erik Johnston	2b084c5b71	Merge device list replication streams (#14833 )	2023-01-17 09:29:58 +00:00
Erik Johnston	73ff493dfb	Merge account data streams (#14826 )	2023-01-13 14:57:43 +00:00
Dirk Klimpel	8d5325ec0c	Drop unused table `presence` (#14825 )	2023-01-13 14:17:03 +00:00
Richard van der Hoff	0f061f39f0	Merge remote-tracking branch 'origin/release-v1.75' into develop	2023-01-12 16:45:23 +00:00
Erik Johnston	84ce93c12f	Fix race calling `/members?at=` (#14817 ) Fixes #14814	2023-01-12 10:29:09 +00:00
reivilibre	d6bda5addd	Add index to improve performance of the `/timestamp_to_event` endpoint used for jumping to a specific date in the timeline of a room. (#14799 )	2023-01-11 12:29:13 +00:00
reivilibre	ba4ea7d13f	Batch up replication requests to request the resyncing of remote users's devices. (#14716 )	2023-01-10 11:17:59 +00:00
Nick Mills-Barrett	db1cfe9c80	Update all stream IDs after processing replication rows (#14723 ) This creates a new store method, `process_replication_position` that is called after `process_replication_rows`. By moving stream ID advances here this guarantees any relevant cache invalidations will have been applied before the stream is advanced. This avoids race conditions where Python switches between threads mid way through processing the `process_replication_rows` method where stream IDs may be advanced before caches are invalidated due to class resolution ordering. See this comment/issue for further discussion: https://github.com/matrix-org/synapse/issues/14158#issuecomment-1344048703	2023-01-04 11:49:26 +00:00
Andrew Morgan	c4456114e1	Add experimental support for MSC3391: deleting account data (#14714 )	2023-01-01 03:40:46 +00:00
reivilibre	2888d7ec83	Faster remote room joins: invalidate caches and unblock requests when receiving un-partial-stated event notifications over replication. [rei:frrj/streams/unpsr] (#14546 )	2022-12-19 14:57:51 +00:00
reivilibre	fb60cb16fe	Faster remote room joins: stream the un-partial-stating of events over replication. [rei:frrj/streams/unpsr] (#14545 )	2022-12-14 14:47:11 +00:00
Patrick Cloke	24a97b3e71	Delete event_push_summary_unique_index again. (#14669 ) if a Synapse deployment upgraded (from < 1.62.0 to >= 1.70.0) then it is possible for schema deltas to run before background updates causing drift in the database schema due to: 1. A delta registered a background update to create an index. 2. A delta dropped the above index if it exists (but it yet exist won't since the background job hasn't run). 3. The code assumed the index was dropped. To fix this we: 1. Cancel the background update which could create the index. 2. Drop the index again. 3. Drop a related index which is dropped by the background update.	2022-12-14 09:25:33 -05:00
David Robertson	e2a1adbf5d	Allow selecting "prejoin" events by state keys (#14642 ) * Declare new config * Parse new config * Read new config * Don't use trial/our TestCase where it's not needed Before: ``` $ time trial tests/events/test_utils.py > /dev/null real 0m2.277s user 0m2.186s sys 0m0.083s ``` After: ``` $ time trial tests/events/test_utils.py > /dev/null real 0m0.566s user 0m0.508s sys 0m0.056s ``` * Helper to upsert to event fields without exceeding size limits. * Use helper when adding invite/knock state Now that we allow admins to include events in prejoin room state with arbitrary state keys, be a good Matrix citizen and ensure they don't accidentally create an oversized event. * Changelog * Move StateFilter tests should have done this in #14668 * Add extra methods to StateFilter * Use StateFilter * Ensure test file enforces typed defs; alphabetise * Workaround surprising get_current_state_ids * Whoops, fix mypy	2022-12-13 00:54:46 +00:00
David Robertson	3d87847ecc	Enable `--warn-redundant-casts` option in mypy (#14671 ) * Enable `--warn-redundant-casts` option in mypy Doesn't do much but helps me sleep better at night. * Changelog * Fix name of the ignore * Fix one more missed cast Not sure why I didn't see this one locally, maybe I needed a poetry update * Remove old comment Co-authored-by: Patrick Cloke <clokep@users.noreply.github.com> Co-authored-by: Patrick Cloke <clokep@users.noreply.github.com>	2022-12-12 21:25:07 +00:00
David Robertson	b5b5f66084	Move `StateFilter` to `synapse.types` (#14668 ) * Move `StateFilter` to `synapse.types` * Changelog	2022-12-12 16:19:30 +00:00
reivilibre	74b89c2761	Revert the deletion of stale devices due to performance issues. (#14662 )	2022-12-12 13:55:23 +00:00
Brendan Abolivier	2a3cd59dd0	Add optional ICU support for user search (#14464 ) Fixes #13655 This change uses ICU (International Components for Unicode) to improve boundary detection in user search. This change also adds a new dependency on libicu-dev and pkg-config for the Debian packages, which are available in all supported distros.	2022-12-12 13:21:17 +01:00
Sean Quah	373c485d8c	Handle half-created indices in receipts index background update (#14650 ) When Synapse is terminated while running the background update to create the `receipts_graph` or `receipts_linearized` indexes, the indexes may be successfully created (or marked as invalid on postgres) while the background update remains unfinished. When Synapse next starts up, the background update will fail because the index already exists, or exists but is invalid on postgres. Use the existing code to create indices in background updates, since it handles these edge cases. Signed-off-by: Sean Quah <seanq@matrix.org>	2022-12-09 23:02:11 +00:00
Patrick Cloke	3ac412b4e2	Require types in tests.storage. (#14646 ) Adds missing type hints to `tests.storage` package and does not allow untyped definitions.	2022-12-09 12:36:32 -05:00
Erik Johnston	94bc21e69f	Limit the number of devices we delete at once (#14649 )	2022-12-09 13:31:32 +00:00
Erik Johnston	c2de2ca630	Delete stale non-e2e devices for users, take 2 (#14595 ) This should help reduce the number of devices e.g. simple bots the repeatedly login rack up. We only delete non-e2e devices as they should be safe to delete, whereas if we delete e2e devices for a user we may accidentally break their ability to receive e2e keys for a message.	2022-12-09 09:37:07 +00:00
Patrick Cloke	c369e95691	Rebuild the user directory and stats tables. (#14643 ) Due to the various fixes to the StreamChangeCache it is not safe to trust the information in the user directory or room/user stats tables. Rebuild them as background jobs. In particular see `da77720752` (#14639), and `6a8310f3df` (#14435). Maybe also be related to `fac8a38525` (#14592).	2022-12-08 11:40:20 -05:00
reivilibre	cf1059d045	Fix a long-standing bug where the user directory would return 1 more row than requested. (#14631 )	2022-12-07 11:19:43 +00:00
Richard van der Hoff	cb59e08062	Improve logging and opentracing for to-device message handling (#14598 ) A batch of changes intended to make it easier to trace to-device messages through the system. The intention here is that a client can set a property org.matrix.msgid in any to-device message it sends. That ID is then included in any tracing or logging related to the message. (Suggestions as to where this field should be documented welcome. I'm not enthusiastic about speccing it - it's very much an optional extra to help with debugging.) I've also generally improved the data we send to opentracing for these messages.	2022-12-06 09:52:55 +00:00
Erik Johnston	cee9445884	Better return type for `get_all_entities_changed` (#14604 ) Help callers from using the return value incorrectly by ensuring that callers explicitly check if there was a cache hit or not.	2022-12-05 15:19:14 -05:00
reivilibre	501f62d1a6	Faster remote room joins: stream the un-partial-stating of rooms over replication. [rei:frrj/streams/unpsr] (#14473 )	2022-12-05 13:07:55 +00:00
Patrick Cloke	fac8a38525	Properly handle unknown results for the stream change cache. (#14592 ) StreamChangeCache.get_all_changed_entities can return None to signify it does not have information at the given stream position. Two callers (related to device lists and presence) were treating this response the same as an empty list (i.e. there being no updates).	2022-12-02 10:28:41 -05:00
David Robertson	781b14ec69	Merge branch 'release-v1.73' into develop	2022-12-01 13:43:30 +00:00
Nick Mills-Barrett	e8bce8999f	Aggregate unread notif count query for badge count calculation (#14255 ) Fetch the unread notification counts used by the badge counts in push notifications for all rooms at once (instead of fetching them per room).	2022-11-30 08:45:06 -05:00
David Robertson	c29e2c6306	Revert "POC delete stale non-e2e devices for users (#14038 )" (#14582 )	2022-11-29 17:48:48 +00:00
David Robertson	e860316818	Fix `UndefinedColumn: column "key_json" does not exist` errors when handling users with more than 50 non-E2E devices (#14580 )	2022-11-29 13:05:07 +00:00
Erik Johnston	c7e29ca277	POC delete stale non-e2e devices for users (#14038 ) This should help reduce the number of devices e.g. simple bots the repeatedly login rack up. We only delete non-e2e devices as they should be safe to delete, whereas if we delete e2e devices for a user we may accidentally break their ability to receive e2e keys for a message. Co-authored-by: Patrick Cloke <clokep@users.noreply.github.com> Co-authored-by: Sean Quah <8349537+squahtx@users.noreply.github.com>	2022-11-29 10:36:41 +00:00
Travis Ralston	9ccc09fe9e	Support MSC1767's `content.body` behaviour; Add base rules from MSC3933 (#14524 ) * Support MSC1767's `content.body` behaviour in push rules * Add the base rules from MSC3933 * Changelog entry * Flip condition around for finding `m.markup` * Remove forgotten import	2022-11-28 18:02:41 -07:00
Andrew Ferrazzutti	1183c372fa	Use `device_one_time_keys_count` to match MSC3202 (#14565 ) * Use `device_one_time_keys_count` to match MSC3202 Rename the `device_one_time_key_counts` key in responses to `device_one_time_keys_count` to match the name specified by MSC3202. Also change related variable/class names for consistency. Signed-off-by: Andrew Ferrazzutti <andrewf@element.io> * Update changelog.d/14565.misc * Revert name change for `one_time_key_counts` key as this is a different key altogether from `device_one_time_keys_count`, which is used for `/sync` instead of appservice transactions. Signed-off-by: Andrew Ferrazzutti <andrewf@element.io>	2022-11-28 16:17:29 +00:00
Sean Quah	f792dd74e1	Remove option to skip locking of tables during emulated upserts (#14469 ) To perform an emulated upsert into a table safely, we must either: * lock the table, * be the only writer upserting into the table * or rely on another unique index being present. When the 2nd or 3rd cases were applicable, we previously avoided locking the table as an optimization. However, as seen in #14406, it is easy to slip up when adding new schema deltas and corrupt the database. The only time we lock when performing emulated upserts is while waiting for background updates on postgres. On sqlite, we do no locking at all. Let's remove the option to skip locking tables, so that we don't shoot ourselves in the foot again. Signed-off-by: Sean Quah <seanq@matrix.org>	2022-11-28 13:42:06 +00:00

... 2 3 4 5 6 ...

5083 Commits