zulip

Commit Graph

Author	SHA1	Message	Date
Tim Abbott	2c7f9ce0fc	puppet: Fix puppet-lint warnings in various manifests. Apparently, `puppet-lint` on Ubuntu trusty throws warnings for certain quoting patterns that are OK in modern `puppet-lint`. I believe the old Zulip code was actually correct (i.e. the old `puppet-lint` implementation was the problem), but it seems worth changing anyway to suppress the warnings. We also exclude more of puppet-apt from linting, since it's third-party code.	2018-08-28 13:46:31 -07:00
Tim Abbott	b53a712856	nginx: Update configuration for using certbot certs everywhere.	2018-08-22 11:59:15 -07:00
Tim Abbott	90828297e4	puppet-lint: Enforce double_quoted_strings check. This makes our puppet codebase more consistent by using single-quoted strings consistently.	2018-08-13 12:31:19 -07:00
Tim Abbott	d0b51b70f4	puppet-lint: Enforce 2sp_soft_tables puppet-lint check. This cleans up the puppet codebase's whitespace formatting to be more consistent.	2018-08-13 12:31:16 -07:00
Tim Abbott	b26e0a957d	puppet-lint: Enforce arrow_alignment check. This fixes all exceptions in our puppet codebase to this lint rule.	2018-08-13 12:30:57 -07:00
Aditya Bansal	710d4507de	puppet-lint: Fix lines longer than 140 characters lint warnings. We fix these by adding ignore statements in a bunch of files where this error popped up. We target only specific lines using the ignore statements and not the entire files.	2018-08-07 10:03:40 -07:00
Anders Kaseorg	edfd5ef992	setup_disks.sh: Fix shellcheck warnings. In puppet/zulip_ops/files/postgresql/setup_disks.sh line 15: array_name=$(mdadm --examine --scan \| sed 's/.*name=//') ^-- SC2034: array_name appears unused. Verify use (or export if used externally). Signed-off-by: Anders Kaseorg <andersk@mit.edu>	2018-08-03 09:15:26 -07:00
Anders Kaseorg	5a0fecc2d5	munin_plugins: Fix shellcheck warnings. In puppet/zulip_ops/files/munin-plugins/rabbitmq_connections line 66: echo "connections.value $(HOME=$HOME rabbitmqctl list_connections \| grep -v "^Listing" \| grep -v "done.$" \| wc -l)" ^-- SC2126: Consider using grep -c instead of grep\|wc -l. In puppet/zulip_ops/files/munin-plugins/rabbitmq_consumers line 32: VHOST=${vhost:-"/"} ^-- SC2034: VHOST appears unused. Verify use (or export if used externally). In puppet/zulip_ops/files/munin-plugins/rabbitmq_messages line 32: VHOST=${vhost:-"/"} ^-- SC2034: VHOST appears unused. Verify use (or export if used externally). In puppet/zulip_ops/files/munin-plugins/rabbitmq_messages_unacknowledged line 32: VHOST=${vhost:-"/"} ^-- SC2034: VHOST appears unused. Verify use (or export if used externally). In puppet/zulip_ops/files/munin-plugins/rabbitmq_messages_uncommitted line 32: VHOST=${vhost:-"/"} ^-- SC2034: VHOST appears unused. Verify use (or export if used externally). In puppet/zulip_ops/files/munin-plugins/rabbitmq_queue_memory line 32: VHOST=${vhost:-"/"} ^-- SC2034: VHOST appears unused. Verify use (or export if used externally). Signed-off-by: Anders Kaseorg <andersk@mit.edu>	2018-08-03 09:15:08 -07:00
Tim Abbott	02ae71f27f	api: Stop using API keys for Django->Tornado authentication. As part of our effort to change the data model away from each user having a single API key, we're eliminating the couple requests that were made from Django to Tornado (as part of a /register or home request) where we used the user's API key grabbed from the database for authentication. Instead, we use the (already existing) internal_notify_view authentication mechanism, which uses the SHARED_SECRET setting for security, for these requests, and just fetch the user object using get_user_profile_by_id directly. Tweaked by Yago to include the new /api/v1/events/internal endpoint in the exempt_patterns list in test_helpers, since it's an endpoint we call through Tornado. Also added a couple missing return type annotations.	2018-07-30 12:28:31 -07:00
Tim Abbott	07af59d4cc	tornado: Split get_events_backend into two functions. The lower-layer function, now called get_events_backend, is intended to be called by multiple code paths (including the upcoming get_events_internal).	2018-07-30 12:28:31 -07:00
Tim Abbott	63fe39e381	zulip_ops: Disable Ubuntu's built-in update-motd.d files. We can't really do this in the zulip manifests (since it's sorta a sysadmin policy decision), but these scripts can cause significant load when Nagios logs into a server (because many of them take 50ms or more of work to run). So we just get rid of them.	2018-05-06 18:47:40 -07:00
Tim Abbott	4e8487c886	nagios: Bump maximum processes limits. These seemed to be flapping for no good reason.	2018-05-02 11:12:47 -07:00
Tim Abbott	718492638b	puppet: Fix name for dhcpcd5 package. Apparently the name dhcpcd isn't installable.	2018-04-23 11:32:07 -07:00
Tim Abbott	35aa4f0377	puppet: Sort ensure attributes to be always first. This inconsistency was flagged by puppet-lint.	2018-04-22 23:41:49 -07:00
Tim Abbott	e103c2ff2d	puppet: Switch to modern quoted, octal file modes. This is one of the prerequisite tasks for Puppet 4 support. Constructed using puppet-lint.	2018-04-22 23:30:48 -07:00
Tim Abbott	62b12e0c34	zulip_ops: Add missing dependency on dhcpcd.	2018-04-19 14:27:48 -07:00
neiljp (Neil Pilgrim)	090b47ed19	mypy: Add explicit Optional for default=None parameters in various files.	2018-03-28 12:31:51 -07:00
neiljp (Neil Pilgrim)	f32f3cbf72	mypy: Amend zulip-ec2-configure-interfaces to avoid None.	2018-03-23 11:39:54 -07:00
Tim Abbott	d98be2f19f	puppet: Only run analytics Nagios checks on machine running cron. Running this on additional machines would be redundant; additionally, the FillState checker cron job runs only on cron systems, so this will crash on other app frontends.	2018-03-06 13:38:27 -08:00
Tim Abbott	8e8faab006	puppet: Move clearsessions cron job to app_frontend_once. While this is a different system than I'd written up in #8004, I think this is a better solution to the general problem of cron jobs to run on just one server. Fixes #8004.	2018-03-06 13:35:51 -08:00
Tim Abbott	3ae645ed12	puppet: Rename analytics.pp to app_frontend_once.pp.	2018-03-06 13:35:51 -08:00
Tim Abbott	24b6106c9c	puppet: Dsiable checking for evictions in memcached nagios. Zulip's caching model for message history is such that it is normal and healthy for there to eventually be a nontrivial volume of evictions.	2018-03-06 13:34:02 -08:00
Greg Price	4475950ddf	queue: Restore prematurely-cut upgrade path. Revert `c8f034e9a` "queue: Remove missedmessage_email_senders code." As the comment in the code says, it ensures a smooth upgrade path from 1.7.x; we can delete it in master after 1.8.0 is released. The removal commit was merged early due to a communication failure.	2018-02-28 11:15:53 -08:00
Umair Khan	c8f034e9a0	queue: Remove missedmessage_email_senders code. After `68513952fb`, all emails are sent through email_senders queue. This commit removes code related to the legacy queue.	2018-02-21 16:43:56 -08:00
Tim Abbott	005b0fb566	puppet: Clean up ssh authorized_keys configuration rules.	2018-02-09 16:37:03 -08:00
Tim Abbott	aca25b6f0a	puppet: Move ssh configuration to use notify. This handles more correctly the case where we're using the upstream sshd_config file.	2018-02-09 16:37:03 -08:00
Tim Abbott	486de8abfc	puppet: Edit some rules to support chat.zulip.org. This should make it possible to use the zulip_ops base rules successfully on chat.zulip.org. Many of the changes in this commit are hacks and probably can be cleaned up later, but given that we plan to drop trusty support soon, it's likely that most of them will simply be deleted then.	2018-02-09 16:37:03 -08:00
Rishi Gupta	1d581a9c6e	nagios: Add nagios check for analytics state. This should help us detect issues where the analytics cron jobs aren't running properly. The cron/nagios part of the implementation done by tabbott.	2018-02-09 16:36:05 -08:00
Tim Abbott	9ed2a94b8c	nagios: Add configuration designed for full-stack servers. This doesn't yet pass all Nagios checks correctly, and still has a few flaws: * The ideal setup code for the `nagios` user in the database isn't included. * Some of the other details are a bit off; we need to split some host roles. But it's better than nothing, and we can iterate from here.	2018-01-24 14:16:03 -08:00
Tim Abbott	2365b13b68	puppet: Move postgres Nagios plugin to main postgres-common. This plugins package is required in order to use Nagios checks to verify the Zulip postgres database, and thus belongs in the default package set.	2018-01-23 10:31:48 -08:00
Umair Khan	68513952fb	email-worker: Create EmailSendingWorker. This commit just copies all the code from MissedMessageSendingWorker class to a new EmailSendingWorker class. All the logic to send an email through a queue was already there. This commit only makes the logic generic. It does so by creating a special purpose queue called 'email_senders' to send any type of email. To make MissedMessageSendingWorker still work we derive it from EmailSendingWorker. All the tests that were testing MissedMessageSendingWorker now run against EmailSendingWorker.	2017-12-20 19:36:27 -08:00
Vishnu Ks	766511e519	actions: Mark all messages as read when user unsubscribes from stream. This fixes a bug where, when a user is unsubscribed from a stream, they might have unread messages on that stream leak. While it might seem to be a minor problem, it can cause significant problems for computing the `unread_msgs` data structures, since it means we need to add an extra filter for whether the user is still subscribed, either in the backend or in the UI. Fixes #7095.	2017-11-21 20:09:17 -08:00
Tim Abbott	94554c65da	certbot: Modify nginx configuration to support automated renewal.	2017-11-08 12:32:26 -08:00
Tim Abbott	62bb465896	puppet: Modify lb0 nginx configuration.	2017-11-08 12:32:26 -08:00
rht	549a26860f	refactor: Remove six.moves.range import.	2017-11-07 10:46:42 -08:00
Tim Abbott	0d1194811f	mypy: Remove ignores for a few typeshed bugs fixed upstream.	2017-10-27 17:09:00 -07:00
Tim Abbott	540cae19a8	puppet: Remove obsolete sparkle configuration. Sparkle was the auto-update system used by the legacy desktop app. We haven't been capable of using it for auto-update in years, so there's no reason to keep around the configuration. The new Electron app uses a different system anyway.	2017-10-19 16:35:55 -07:00
rht	b57289aacd	py3: Remove all `from __future__ import print_function. Except for these files: - tools/linter_lib/* - tools/lib - tools/lister.py	2017-10-18 12:07:19 -07:00
rht	2f3ae84e5a	py3: Remove all `__future__ import division`.	2017-10-17 23:09:12 -07:00
Tim Abbott	6a5cb0e48c	puppet: Make problems with Zephyr mirroring pageable. Generally this indicates sending messages is completely broken.	2017-10-12 00:16:32 -07:00
rht	de30400fc5	pg_backup_and_purge.py: Remove .py extension.	2017-10-08 15:32:43 -07:00
Tim Abbott	47c5aae5b2	log2zulip: Enforce using python 3 in cron job. We aren't guaranteed to have the Zulip dependencies installed on Python 2.	2017-10-06 16:37:17 -07:00
Tim Abbott	f2055397c1	nagios: Update apache configuration to be generated. Since this is basically just stock Apache configuration for Nagios with a hostname put in, we can just fetch the hostname from our configuration.	2017-10-05 21:51:29 -07:00
Tim Abbott	3af01bed85	puppet: Simplify zulip_ops nginx configuration. Whatever dist/ functionality this had in 2014 is now served by zulip.org, and since this serves as a sample, it should be as simple as possible. Previously, this was more cluttered than it needed to be.	2017-10-05 21:17:57 -07:00
Tim Abbott	e6e7bcf6e1	nagios: Move camo_check_url into configuration.	2017-10-05 21:09:24 -07:00
Tim Abbott	1c453fdf2a	puppet: Add redis_password file for Nagios. This allows the Nagios user to access redis without having full access to the redis system. Ideally, this would eventually use a password that only has statistics read access, but I'm not sure redis supports that.	2017-10-05 20:42:07 -07:00
Tim Abbott	13a36d9af3	puppet: Make old redis_tunnel configuration usable. This old puppet configuration was never really used, and regardless hardcoded an ancient zulip.net hostname. We fix this to use the zulipconf system to get the host domain (though not, at present, the hostname).	2017-10-05 20:40:22 -07:00
Tim Abbott	96c3014da0	nagios: Automate configuration of outgoing email with msmtp. Now we no longer need to check in a bunch of hostnames in order to configure Nagios.	2017-10-05 20:29:47 -07:00
Tim Abbott	5b4c260c3f	puppet: Add munin apache auth configuration. This is completely stock configuration, and seems to be required for munin to run properly.	2017-10-05 20:17:12 -07:00
Tim Abbott	ba7be4102e	puppet: Update munin tunnels configuration to use zulipconf. This eliminates another old hardcoding of zulip.net.	2017-10-05 20:14:43 -07:00
Tim Abbott	162eaf8917	nagios: Modify check for swap to allow no swap. If a machine is configured with no swap intentationally, that shouldn't be a Nagios problem. This alert is intended to flag machines which are swapping.	2017-10-05 20:07:44 -07:00
Tim Abbott	80a16bf873	nagios: Fix path to source zulip_nagios.cfg. Arguably, we should make this a symlink, but it's probably a good idea to have every change in the production Nagios configuration go through the zulip-puppet-apply diff experience.	2017-10-05 20:06:50 -07:00
Tim Abbott	886a8853ac	nagios: Move server-specific config into hostgroups. These new hostgroups exist so we can eliminate explicit references to individual hosts in services.cfg.	2017-10-05 20:06:48 -07:00
Tim Abbott	b6ce9583a9	nagios: Fetch list of hosts from zulip.conf. This makes this much more configurable and much less hardcoded.	2017-10-05 20:06:30 -07:00
Tim Abbott	5193936bc3	nagios: Add Memcached and Redis monitoring. These are standard Nagios plugins that might be sometimes helpful.	2017-10-05 20:06:16 -07:00
Tim Abbott	f7d554d533	nagios: Rename zmirror2 to zmirrorp in configuration. The "p" stands for "personals", aka zephyr private messages, which is what this host manages.	2017-10-05 20:06:08 -07:00
Tim Abbott	062d280914	puppet: Clean up unnecessary pagerduty_nagios.cfg.	2017-10-05 19:23:33 -07:00
Tim Abbott	7e328ba865	nagios: Move email addresses for contacts into variables.	2017-10-05 19:23:33 -07:00
Tim Abbott	6017d3dec5	puppet: Move contacts.cfg to be a template.	2017-10-05 19:23:33 -07:00
Tim Abbott	09aec3e467	puppet: Move hosts.cfg to be managed by a template.	2017-10-05 19:23:33 -07:00
Tim Abbott	692f4b77d1	puppet: Remove messy Nagios crontab.	2017-10-05 19:23:33 -07:00
Tim Abbott	26982ff55f	puppet: Remove pageduty_nagios.pl. This hasn't been used in like 4 years, and clutters the repo.	2017-10-05 18:46:09 -07:00
Tim Abbott	5a80c029a2	nagios: Update path to sync_public_streams to match new config.	2017-10-05 13:34:27 -07:00
Tim Abbott	fdd021fd6a	zephyr-mirror: Update supervisor configuration for repository split. This now points to the path of the integration in the new package.	2017-10-05 13:18:37 -07:00
Tim Abbott	1eff717146	zephyr-mirror: Update cron job to use python-zulip-api. This is a deferred follow-up project to the repository split.	2017-10-05 13:07:45 -07:00
Greg Price	d02101a401	APNs: Rip out the existing, broken implementation. This code empirically doesn't work. It's not entirely clear why, even having done quite a bit of debugging; partly because the code is quite convoluted, and because it shows the symptoms of people making changes over time without really understanding how it was supposed to work. Moreover, this code targets an old version of the APNs provider API. Apple deprecated that in 2015, in favor of a shiny new one which uses HTTP/2 to meet the same needs for concurrency and scale that the old one had to do a bunch of ad-hoc protocol design for. So, rip this code out. We'll build a pathway to the new API from scratch; it's not that complicated.	2017-08-26 14:16:05 -07:00
Greg Price	a099e698e2	py3: Switch almost all shebang lines to use `python3`. This causes `upgrade-zulip-from-git`, as well as a no-option run of `tools/build-release-tarball`, to produce a Zulip install running Python 3, rather than Python 2. In particular this means that the virtualenv we create, in which all application code runs, is Python 3. One shebang line, on `zulip-ec2-configure-interfaces`, explicitly keeps Python 2, and at least one external ops script, `wal-e`, also still runs on Python 2. See discussion on the respective previous commits that made those explicit. There may also be some other third-party scripts we use, outside of this source tree and running outside our virtualenv, that still run on Python 2.	2017-08-16 17:54:43 -07:00
Greg Price	e4d1d22e9f	py3: Explicitly keep our wal-e PostgreSQL replication on Python 2. On `trusty` there is no package for `boto` or `gevent` on Python 3, both of which are dependencies of `wal-e` (at the version we've pinned.) This is something used only on database servers and only in a replication scenario, and it doesn't involve any of our code outside the wal-e repo, so the Python version it uses is quite independent of the Zulip application server itself and the rest of our code. For now, keep it explicitly on Python 2 while we move forward for most everything else.	2017-08-15 17:30:31 -07:00
Greg Price	2a4d851a7c	py3: Explicitly keep one boto-using ops script on Python 2. This script in `zulip_ops` is handy for managing EC2 instances. It uses `boto`, which isn't available in `trusty` for Python 3. The use of `boto` here isn't particularly deep, so we could replace it with some more manual HTTP calls if it comes to that. For now, just mark it to stay on Python 2 while we move the app and all the rest of the ops code (except this and another straggler or two) to Python 3. Also make a comment on this package in the Puppet manifest clearer about what it specifically refers to.	2017-08-15 17:30:31 -07:00
Greg Price	61666a9262	zulip_ops: Delete the long-disused `stats1.zulip.net` config and its dependencies. This consists of the `zulip_ops::stats` Puppet class, which has apparently not been used since 2014, and a number of files that I believe were only used for that. Also a couple of tiny loose ends in other files.	2017-08-15 17:30:31 -07:00
Greg Price	0042fc0c19	py3: Move `python-gevent` dependency to narrow its scope. This is only actually used in our `wal-e` setup, which is in zulip_ops::postgres_common. (In fact the only mentions of `gevent` in our whole Git history are for `wal-e`.) So remove where we mention it on the broader zulip::postgres_common module, and move it where it's needed. This follows up on `98cef0ab4` by eliminating the only dependency outside of the `zulip_ops` Puppet tree on a system Python-library package which isn't available in `trusty` for Python 3.	2017-08-15 17:30:31 -07:00
Greg Price	98cef0ab48	py3: Augment all mentions of system Python packages to include Python 3. In some of these contexts, we may still be using the Python 2 version, but at least this should eliminate running into `ImportError`s one by one in scripts that run outside a virtualenv, as we update their shebangs to refer to Python 3. Several Python libraries we use don't come in Python 3 versions on trusty: gevent, boto, twisted, django, django-tagging, whisper. The latter two don't come in Python 3 versions even on xenial. So some work required before we can actually switch the code that relies on those libraries to run as Python 3 -- probably the best solution will be to backport them all in our apt repo. (All but `whisper` are packaged in zesty; `whisper` upstream just grew Python 3 support this year.)	2017-08-09 14:07:05 -07:00
Greg Price	b8089bdd1c	api: Update log2zulip cron job to find the script at its new path.	2017-07-31 21:24:02 -07:00
Greg Price	c127630dcf	Delete some obsolete usage-stats tools. These are no longer useful, with our spiffy new analytics framework, and we haven't in fact been using them for some time, while the `active-user-stats` cron job does cause regular mail from cron. Just delete them.	2017-07-31 17:06:15 -07:00
neiljp (Neil Pilgrim)	8cabce9f5e	mypy: For EC2, Ensure to_configure is passed a not-None argument.	2017-07-17 16:57:42 -07:00
neiljp (Neil Pilgrim)	ba51958c40	mypy: For EC2, pre-assign address & gateway to enable assertion.	2017-07-17 16:57:42 -07:00
neiljp (Neil Pilgrim)	fd941e8f88	mypy: For EC2, make guess_gateway return None if address is None.	2017-07-17 16:57:42 -07:00
Aditya Bansal	5989c88545	pep8: Make compliant zulip-ec2-configure-interfaces with rule E261.	2017-05-31 17:07:15 -07:00
Aditya Bansal	6b8e85e065	pep8: Make compliant check_zephyr_mirror with rule E261.	2017-05-31 17:07:15 -07:00
Aditya Bansal	49ae51f23a	pep8: Make compliant check_user_zephyr_mirror_liveness with rule E261.	2017-05-31 17:07:15 -07:00
Aditya Bansal	6d0927ed0b	pep8: Add compliance with rule E261 to check_personal_zephyr_mirrors.	2017-05-31 17:07:15 -07:00
Tim Abbott	c5bc1265f3	bots: Move log2zulip into api/integrations.	2017-05-26 15:15:56 -07:00
Tim Abbott	55f69c677a	puppet: Remove obsolete zuliprc.nagios file. This hasn't done anything for years.	2017-05-26 15:14:12 -07:00
Reid Barton	ccb4c5c26f	bots: Move zephyr-related files to api/integrations/zephyr/.	2017-05-26 15:07:02 -07:00
Elliott Jin	0ec9e54954	bots: Add queue and QueueProcessingWorker for embedded bots.	2017-05-25 15:00:51 -07:00
Tim Abbott	1d1b0894a3	zulip_ops: Add logrotate configuration from main zulip.	2017-05-15 21:54:35 -07:00
Tim Abbott	e049ea01b1	puppet: Update munin configuration to work with modern munin.	2017-05-15 21:49:53 -07:00
Tim Abbott	6871c227cd	puppet: Add missing restarts of Nagios on config updates.	2017-05-15 21:49:53 -07:00
vaibhav	8881b5eb9f	Outgoing Webhook System: Check for @-mentioned outgoing webhook bots. Also puts them into a processing queue, though the queue processor does nothing. Rewritten by tabbott to avoid unnecessary database queries in do_send_messages.	2017-05-02 09:22:04 -07:00
hackerkid	b2504084ab	Replace timezone.now with timezone_now.	2017-04-16 12:28:56 -07:00
Tim Abbott	6a5e98b77e	puppet: Increase MaxStartups SSH configuration.	2017-03-08 22:28:16 -08:00
K.Kanakhin	6a801db1c2	missed-emails-sending: Move email sending to separate queue worker. - Add new 'missedmessage_email_senders' queue for sending missed messages emails. - Add the new worker to process 'missedmessage_email_senders' queue. - Split aggregation missed messages and sending missed messages email to separate queue workers. - Adapt tests for sending missed emails to the new logic. Fixes #2607	2017-03-07 20:08:40 -08:00
Raghav Jajodia	a3a03bd6a5	mypy: Added Dict, List and Set imports. Fixed mypy errors associated with the upgrade.	2017-03-04 14:33:44 -08:00
Rishi Gupta	3d07ac0c49	Change timezone-naive datetimes to use timezone.now() where safe to do so. Change timezone-naive datetimes to use timezone.now() in cases where there is no change in behavior.	2017-03-01 22:54:28 -08:00
Tim Abbott	fa8045a484	puppet: Add websockets Nagios test to configuration. Since browser clients send messages via websockets and not the API, this is an important element in making sure mission-critical Zulip functionality is working.	2017-02-08 11:13:19 -08:00
Tim Abbott	ba5f454be5	puppet: Extract zulip::analytics. I'm not altogether happy with this (a better solution would be database-level locking), but I think it solves the immediate problem of folks with 2 servers being very likely to run analytics on both of them.	2017-02-07 12:29:15 -08:00
Tim Abbott	36d54cf5ff	Replace references to zulip.com/dist with zulip.org/dist. Now that zulip.org has all the files to distribute, there's no reason to still point to the soon-to-be-decommissioned zulip.com/dist.	2017-01-28 17:56:25 -08:00
Tim Abbott	4e171ce787	lint: Clean up E126 PEP-8 rule.	2017-01-23 22:06:13 -08:00
Tim Abbott	d6e38e2a5c	lint: Clean up E123 PEP-8 rule.	2017-01-23 21:34:26 -08:00
JefftheBest1	9de75f5167	Fixed typos with separate	2017-01-12 04:52:05 -08:00
JefftheBest1	ff8639f9db	Fixed typos with threshold.	2017-01-12 04:50:20 -08:00
JefftheBest1	5008f45112	Fixed typo in munin.conf.erb	2017-01-12 04:49:19 -08:00
Tim Abbott	3e32102016	nagios: Fix various critical issues not tagged as pageable.	2017-01-06 21:49:20 -08:00
Tim Abbott	edebf7619b	puppet: Add PAM common_session disabling systemd-login. This fixes a weird problem with systemd where logging into a server via ssh frequently has a 15s+ lag.	2017-01-06 21:49:15 -08:00
Tim Abbott	93c2c19775	nagios: Increase process count limits.	2017-01-06 21:49:15 -08:00
Tim Abbott	2c6cb37385	munin: Add default munin configuration template.	2017-01-06 21:44:57 -08:00
Tim Abbott	9ab8e7ba34	nagios: Disable swap checks for servers with no swap.	2017-01-06 21:39:07 -08:00
Tim Abbott	3e01ed1f73	nagios: Increase NTP max_check_attempts. NTP often suffers from brief interruptions of service that lead to spurious Nagios alerts; it makes sense to suppress these.	2017-01-06 21:32:43 -08:00
Tim Abbott	e4420b08d2	zulip_ops: Disable unattended upgrades of security packages. Since Zulip does not handle e.g. postgres server restarts gracefully, it's best for a system administrator to manually trigger security updates.	2017-01-06 21:30:56 -08:00
Tim Abbott	6f9c73d0e5	zmirror: Update Debathena release in configuration. The zulip_ops configuration is now for xenial, not obsolete wheezy.	2017-01-06 21:30:41 -08:00
Tim Abbott	bd9176d1d9	nagios: Remove some default files. Nagios ships with a bunch of default configuration files that one needs to delete in order to configure it.	2017-01-06 21:25:12 -08:00
Tim Abbott	7083899e77	zulip_ops: Add postgres config for enabling Nagios. The old zulip_ops Nagios configuration depended on Nagios having the ability to login as the zulip user (with essentially full write access); this configuration is helpful for limiting nagios to special "nagios" user with more limited credentials.	2017-01-06 21:24:24 -08:00
Tim Abbott	204edb0f85	zulip_ops: Cleanup pg_hba.conf configuration.	2017-01-06 21:23:51 -08:00
Tim Abbott	30c57eb2ae	zulip_ops: Add basic .emacs for production.	2017-01-06 21:20:21 -08:00
Tim Abbott	eb87d04168	puppet: Remove xxxxx password hardcoding in recovery.conf.	2017-01-06 21:20:21 -08:00
Tim Abbott	6404a1a5ff	zulip_ops: Add nagios-plugins-contrib. This has a number of useful nagios plugins.	2017-01-06 21:19:59 -08:00
Tim Abbott	f7b77008ef	zulip_ops: Add aptitude dependency. This is useful for `aptitude why`.	2017-01-06 21:19:50 -08:00
Tim Abbott	2510a51a8a	zulip_ops: Add letsencrypt dependency.	2017-01-06 21:19:31 -08:00
Tim Abbott	65774e1c4f	zulip_ops: use check_postgres package from apt.	2017-01-06 21:18:55 -08:00
K.Kanakhin	0d8c18a6dd	nagios-plugins: Add websocket checking to nagios message sending test. - Add websocket client to create connection with SockJS websocket server. It contains callback method to launch after connection setup. - Add '--websocket' parameter to 'check_send_receive_time' script to check websocket connection. - Add testing websocket connection to production installation checking. - Add cronjob to launch websocket connection nagios test. This makes it possible for Zulip Nagios monitoring to check for problems impacting the websockets sending code path, which is what all web users use.	2016-12-30 15:36:37 -08:00
Jason Le	144d82305d	mypy: Annotate puppet/zulip_ops.	2016-12-03 11:00:25 -08:00
bulat22101	adebc75740	pep8: Fix E502 violations	2016-12-03 10:56:36 -08:00
Sidhant Bhavnani	8c0c12c1d9	pep8: Fix E303 violations.	2016-12-02 15:34:11 -08:00
Rafid Aslam	c5316b4002	lint: Fix E127 pep8 violations. Fix pep8: E127 continuation line over-indented for visual indent style issue.	2016-12-01 10:23:55 -08:00
Rafid Aslam	7a2282986a	pep8: Fix E225 pep8 violations.	2016-11-28 15:21:15 -08:00
Anders Kaseorg	207cf6302b	Always start python via shebang lines. This is preparation for supporting using Python 3 in production. Signed-off-by: Anders Kaseorg <andersk@mit.edu>	2016-11-26 14:46:37 -08:00
Tim Abbott	2e65dc1206	puppet: make check_send_receive_time target host configurable.	2016-11-02 23:40:53 -07:00
Tim Abbott	eceaf36001	setup_disks: Fix postgres RAID setup to work correctly on Xenial. Nobody's going to run this on Wheezy again.	2016-10-28 11:04:08 -07:00
Tim Abbott	9b7a3f040c	Remove now-unused /json/get_events endpoint.	2016-10-27 21:34:58 -07:00
Tim Abbott	4fbe201187	puppet: Automate autossh process monitoring maintenance. Previously, the Zulip Nagios configuration effectively hardcoded the count for how many system should have autossh connections.	2016-10-26 00:49:03 -07:00
Tim Abbott	6bdb10b71b	puppet: Update emacs dependency to emacs-nox metapackage. This way, one doesn't need to keep updating the dependency every time a new major emacs release comes out.	2016-10-26 00:42:22 -07:00
Tim Abbott	11b5d203f7	sshd_config: Increase MaxStartups. This fixes connection problems when using the full Zulip recommended Nagios configuration against a given server.	2016-10-26 00:41:03 -07:00
Tim Abbott	73f54dd0cb	sshd_config: Add updates from Xenial upstream. It seems worth updating this to match the Linux distro this configuration targets.	2016-10-26 00:40:44 -07:00
Tim Abbott	0a5a2c4eda	nagios: Automate authorized users list maintenance.	2016-10-26 00:37:29 -07:00
Tim Abbott	fa4998db59	puppet: Add zulip_zephyr_mirror plugins.	2016-10-26 00:35:57 -07:00
Tim Abbott	ac4f28050c	zmirror: Remove unnecessary krb5-clients dependency. I'm pretty sure krb5-clients isn't needed to run the Zephyr mirroring service.	2016-10-26 00:35:11 -07:00
Tim Abbott	d490e83645	puppet: Upgrade nagios cgi.cfg with modern defaults.	2016-10-26 00:31:41 -07:00
Tim Abbott	1159ad4857	puppet: Upgrade nagios.cfg with modern defaults.	2016-10-26 00:31:41 -07:00
Tim Abbott	73178e5e5a	puppet: Run check_send_receive_time via a cron job. This allows the actual nagios work involved with check_send_receive_time nagios checks to be done by an unprivileged "nagios" user rather than the "zulip" user.	2016-10-26 00:26:52 -07:00
Tim Abbott	96cf330649	puppet: ssh as the nagios user instead of zulip user. This is a follow-up to `4f58fef54b`, touching services.cfg instead of commands.cfg.	2016-10-26 00:23:47 -07:00
Tim Abbott	a350d43683	puppet: Add recovery.conf configuration to postgres_slave.pp. This file is needed to run a valid postgres slave; it's not clear why this wasn't installed in the original zulip.com configuration.	2016-10-26 00:22:57 -07:00
Tim Abbott	c3727c9886	nagios: Remove old zulip.com trac/git/replica servers. These are unlikely to be relevant to anyone.	2016-10-26 00:21:53 -07:00
Tim Abbott	383f39b543	nagios: Enable allow_empty_hostgroup_assignment. This fixes the configuration being broken when we remove some of the old zulip.com hosts that are unlikely to be of interest to anyone.	2016-10-26 00:19:21 -07:00
Tim Abbott	4f58fef54b	zulip_ops: Use nagios user for all Nagios checks. There's no reason these Nagios checks needs to run as the semi-priviliged Zulip user.	2016-10-26 00:17:26 -07:00
Tim Abbott	32d244dbe5	puppet: Add Nagios checks for other consumers.	2016-10-26 00:11:08 -07:00
Tim Abbott	f1fa4397f3	puppet: Fix package deps for zulip-ec2-configure-interfaces.	2016-10-26 00:11:08 -07:00
Tim Abbott	3448ab4c7a	zulip-ec2-configure: Fix network IDs for Ubuntu Xenial.	2016-10-26 00:11:08 -07:00
Tim Abbott	080dd8c987	nagios: Ignore kthreads in check_procs tests. Modern Linux can have a lot of kernel threads not doing anything. Since this isn't interesting from a monitoring perpsective, we ignore these.	2016-10-26 00:10:40 -07:00
Tim Abbott	4c9a283542	puppet: Remove configuration for old builder host. I don't think this configuration was ever even used; it's just clutter.	2016-10-26 00:01:52 -07:00
Tim Abbott	f9ad75f58e	puppet: Remove configuration for old zulip.com bots host. This configuration didn't do anything anyway and just clutters the repo.	2016-10-26 00:01:29 -07:00

1 2 3 4 5 ...

264 Commits