zulip

Commit Graph

Author	SHA1	Message	Date
Anders Kaseorg	98ed6248e3	apt-repos: Remove now-unneeded Ubuntu 21.10 repository on 22.04. Followup to commit `f8957863a2` (#22055). Signed-off-by: Anders Kaseorg <anders@zulip.com>	2022-05-25 17:25:23 -07:00
Alex Vandiver	6337f17923	upgrade: Add --skip-restart which preps but does not restart. This adds a --skip-restart which makes `deployments/next` in a state where it can be restarted into, but holds off on conducting that restart. This requires many of the same guarantees as `--skip-tornado`, in terms of there being no Puppet or database schema changes between the versions. Enforce those with `--skip-restart`, and also broaden both flags to prevent other, less common changes which nonetheless potentially might affect the other deploy.	2022-05-22 15:07:37 -07:00
Alex Vandiver	86a4e64726	upgrade: Enforce that --skip-tornado does not have Puppet or DB changes.	2022-05-22 15:07:18 -07:00
Alex Vandiver	ef7c2ea0ea	upgrade: Copy cache prefix with --skip-tornado. Because Tornado and Django use memcached as a shared cache for checking session information, they must agree on the prefix used to store those values. Subsequent commits will work to ensure that it is always _safe_ to share that cache.	2022-05-22 14:52:38 -07:00
Alex Vandiver	fa77be6e6c	upgrade: Only run Django system checks once, explicitly. These are expensive, and moving them to one explicit call early has considerable time savings in the critical period: ``` $ hyperfine './manage.py fill_memcached_caches' './manage.py fill_memcached_caches --skip-checks' Benchmark #1: ./manage.py fill_memcached_caches Time (mean ± σ): 5.264 s ± 0.146 s [User: 4.885 s, System: 0.344 s] Range (min … max): 5.119 s … 5.569 s 10 runs Benchmark #2: ./manage.py fill_memcached_caches --skip-checks Time (mean ± σ): 3.090 s ± 0.089 s [User: 2.853 s, System: 0.214 s] Range (min … max): 2.950 s … 3.204 s 10 runs Summary './manage.py fill_memcached_caches --skip-checks' ran 1.70 ± 0.07 times faster than './manage.py fill_memcached_caches' ```	2022-05-22 14:52:38 -07:00
Alex Vandiver	3928606886	restart-server: Treat as a start if nothing is running. Treating the restart as a start is important in reducing the critical period during upgrades -- we call restart even when we suspect the services are stopped, because puppet has a small possibility of placing them in indeterminate state. However, restart orders the workers first, then tornado/django, which prolongs the outage. Recognize when no services are currently started, and switch to acting like a start, not a restart, which places tornado/django first.	2022-05-22 14:52:38 -07:00
Alex Vandiver	3717c329b8	stop-server: Only stop services if they exist and are running. This hides ugly output if the services were already stopped: ``` 2022-03-25 23:26:04,165 upgrade-zulip-stage-2: Stopping Zulip... process-fts-updates: ERROR (not running) zulip-django: ERROR (not running) zulip_deliver_scheduled_emails: ERROR (not running) zulip_deliver_scheduled_messages: ERROR (not running) Zulip stopped successfully! ``` Being able to skip having to shell out to `supervisorctl`, if all services are already stopped is also a significant performance improvement.	2022-05-22 14:52:38 -07:00
Alex Vandiver	2e5a079ef4	upgrade: Check with zulip-puppet-apply to see if we can skip it.	2022-05-22 14:52:38 -07:00
Alex Vandiver	ecfc23bd0b	zulip-puppet-apply: Make --force --noop have an exit code.	2022-05-22 14:52:38 -07:00
Alex Vandiver	c91725bfb5	zulip-puppet-apply: Factor out the --noop returncode logic.	2022-05-22 14:52:38 -07:00
Alex Vandiver	b15d8e0118	upgrade: Skip the pre-work if the server is already stopped. This optimization makes sense if the server is already running, but if it is already stopped, it is just prolonging the downtime.	2022-05-22 14:52:38 -07:00
Alex Vandiver	05af4b0a11	upgrade: Fill caches before the critical period, if possible.	2022-05-22 14:52:38 -07:00
Alex Vandiver	2f7068ffbb	upgrade: Move puppet class renames earlier. These do not need to happen during the critical period when the server is stopped.	2022-05-22 14:52:38 -07:00
Anders Kaseorg	f8957863a2	Revert "apt-repos: Downgrade PostgreSQL to dodge PGroonga regression." This reverts commit `9c8d2b7be3` (#21115). The PostgreSQL fix was released 2022-05-12. Signed-off-by: Anders Kaseorg <anders@zulip.com>	2022-05-17 15:07:37 -07:00
Alex Vandiver	258b658cc0	log-search: Allow multiple search terms. This allows AND'ing multiple terms together.	2022-05-06 17:45:46 -07:00
Alex Vandiver	bd73e7d411	log-search: Factor out argument parsing.	2022-05-06 17:45:46 -07:00
Alex Vandiver	8eab5f6931	log-search: Add status code search. This moves log filename parsing after the filter parsing, as that can now enable --nginx.	2022-05-06 17:45:46 -07:00
Alex Vandiver	0bad002c14	log-search: Factor out logfile name parsing.	2022-05-06 17:45:46 -07:00
Alex Vandiver	67e641f37d	log-search: Add a filter by path.	2022-05-06 17:45:46 -07:00
Alex Vandiver	df47c5a750	log-search: Update docs to include client-id as an option.	2022-05-06 17:45:46 -07:00
Alex Vandiver	b1749259d4	log-search: Fix URLs for non-zulipchat.com hosts.	2022-05-06 17:45:46 -07:00
Alex Vandiver	e3a65b1528	log-search: Some Django log lines do not include hostname.	2022-05-06 17:45:46 -07:00
Alex Vandiver	fe17a4d6d0	log-search: Handle ^C more gracefully.	2022-05-06 17:45:46 -07:00
Alex Vandiver	da4ae3ff24	log-search: Filter out user avatars.	2022-05-06 17:45:46 -07:00
Alex Vandiver	d3ae7480cc	log-search: Handle settings.LOGGING_SHOW_PID.	2022-05-06 17:45:46 -07:00
Alex Vandiver	bd298ba753	log-search: Not all servers are in UTC.	2022-05-06 17:45:46 -07:00
Anders Kaseorg	3cb7d3d1dc	node_cache: Remove node_modules/.cache when copying. Signed-off-by: Anders Kaseorg <anders@zulip.com>	2022-05-04 09:56:07 -07:00
Alex Vandiver	65b99377d2	log-search: Show duration.	2022-05-03 13:44:29 -07:00
Alex Vandiver	056895cc33	log-search: Search for user-ids.	2022-05-03 13:44:29 -07:00
Alex Vandiver	b355a0a63e	log-search: Default to searching python logfiles. These have more accurate timestamps, and have user information -- but are harder to parse, and will not show requests when Django or Tornado is stopped.	2022-05-03 13:44:29 -07:00
Alex Vandiver	ba1237119c	log-search: Add a tool to search nginx logs by IP/hostname. This is a script to search nginx log files by server hostname or client IP address, and output matching lines, all while skipping common and less-interesting request lines.	2022-05-03 13:44:29 -07:00
Alex Vandiver	e13154f089	puppet: Add ksplice support for 22.04.	2022-05-03 12:36:19 -07:00
Alex Vandiver	cda55a40e7	puppet: Add teleport support for 22.04.	2022-05-03 12:36:19 -07:00
Anders Kaseorg	e952641013	install: Resupport Ubuntu 22.04. Signed-off-by: Anders Kaseorg <anders@zulip.com>	2022-05-03 09:41:08 -07:00
Anders Kaseorg	25c87cc7da	zulip-puppet-apply: Work around broken Puppet on Ubuntu 22.04. Signed-off-by: Anders Kaseorg <anders@zulip.com>	2022-05-03 09:41:08 -07:00
Anders Kaseorg	080a806d60	build-pgroonga: Update PGroonga to 2.3.6. Signed-off-by: Anders Kaseorg <anders@zulip.com>	2022-04-29 16:02:45 -07:00
Alex Vandiver	3476f63dca	compare-settings-to-template: Handle prod_settings_template renaming.	2022-04-28 14:52:38 -07:00
Alex Vandiver	b6b6faa404	compare-settings-to-template: Simplify and dedent logic.	2022-04-28 14:52:38 -07:00
Alex Vandiver	d205050ab0	compare-settings-to-template: Fetch 100 per pagination.	2022-04-28 14:52:38 -07:00
Alex Vandiver	d79776f80d	compare-settings-to-template: Paginate through all tags. The default page size is 30, which means this only goes back to 4.6 at present, due to starting with `shared-...` and old `enterprise-...` tags.	2022-04-28 14:52:38 -07:00
Anders Kaseorg	098a514599	python: Use Python 3.8 shlex.join function. Signed-off-by: Anders Kaseorg <anders@zulip.com>	2022-04-27 12:57:49 -07:00
Anders Kaseorg	0451d1e47f	zulip_tools: Replace universal_newlines with text. Generated by pyupgrade. Signed-off-by: Anders Kaseorg <anders@zulip.com>	2022-04-27 12:57:49 -07:00
Anders Kaseorg	a543dcc8e3	Remove Debian 10 support. As a consequence: • Bump minimum supported Python version to 3.8. • Move Vagrant environment to Ubuntu 20.04, which has Python 3.8. • Move CI frontend tests to Ubuntu 20.04. • Move production build test to Ubuntu 20.04. • Move 3.4 upgrade test to Ubuntu 20.04. Signed-off-by: Anders Kaseorg <anders@zulip.com>	2022-04-26 16:32:02 -07:00
Anders Kaseorg	63a1ef0e91	configure-rabbitmq: Remove use of sudo. It already runs as root everywhere except in provision_inner, so move the sudo there. Signed-off-by: Anders Kaseorg <anders@zulip.com>	2022-04-19 12:36:31 -07:00
Anders Kaseorg	cc30ed8ec7	actions: Delete zerver.lib.actions. Signed-off-by: Anders Kaseorg <anders@zulip.com>	2022-04-14 17:14:38 -07:00
Alex Vandiver	09860dc284	check-database-compatibility: Sort and prettify output.	2022-04-06 14:10:46 -07:00
Alex Vandiver	eb31681934	check-database-compatibility: Ignore squashed and renamed migrations. Fixes: #21596.	2022-04-01 16:15:41 -07:00
Alex Vandiver	0af00a3233	upgrade: Mark puppet as having started the server. We previously used restart-server if puppet was run, as a nod to the fact that `supervisor reread && supervisor update` will _start_ service groups that were modified, even if they were previously stopped; this is because they are marked as `autostart=true`, which is honored on service change. However, upgrades want to run while there are no services running. If puppet is run, explicitly set the server as potentially being "up", so that a `shutdown_server()` before migrations, if they exist, will stop services.	2022-03-31 17:21:39 -07:00
Alex Vandiver	e9596637e7	upgrade: Move the shutdown_server calls to where they are relevant. shutdown_server is a noop if the server is already stopped; placing these in each block makes the logic more apparent.	2022-03-31 17:21:39 -07:00
Alex Vandiver	65e19c4fbd	supervisor: 'foo:' also matches 'foo'. `7c4293a7d3` switched to checking if the service was already running, and use `supervisorctl start` if it was not. Unfortunately, `list_supervisor_processes("zulip-tornado:")` did not include `zulip-tornado`, and as such a non-sharded process was always considered to _not_ be running, and was thus started, not restarted. Starting an already-started service is a no-op, and thus non-sharded tornado processes were never restarted. The observed behaviour is that requests to the tornado process attempt to load the user from the cache, with a different prefix from Django, and immediately invalidate the session and eject the user back to the login page. Fix the `list_supervisor_processes` logic to match without the trailing `:*`.	2022-03-31 10:41:41 -07:00

1 2 3 4 5 ...

1244 Commits