zulip

Commit Graph

Author	SHA1	Message	Date
Vishnu KS	51f5701879	export: Canonicalize the email of cross realm bot to default value. Fixes #13496	2020-02-19 14:44:50 -08:00
Mateusz Mandera	920d22524b	import: Use re_map_foreign_keys on the realm column of UserPresence. We forgot to make this adjustment in the recent denormalization of realm into UserPresence. It's needed for imports to work correctly.	2020-02-18 10:45:38 -08:00
Vishnu Ks	5a59bf329e	import: Skip setting user_profile_id metadata only if unavailable.	2020-02-03 14:09:05 -08:00
Vishnu Ks	2ea53a347a	import: Support importing realm icon and logo. Fixes #11216	2020-02-03 14:09:05 -08:00
Ryan Rehman	3dc7d60ffe	muting: Record DateTime when a Topic is muted. This includes the necessary migration to add the date_muted field to the MutedTopic class and populates it with a hard coded value.	2020-02-02 20:49:53 -08:00
Tim Abbott	dd969b5339	install: Remove references to "Zulip Voyager". "Zulip Voyager" was a name invented during the Hack Week to open source Zulip for what a single-system Zulip server might be called, as a Star Trek pun on the code it was based on, "Zulip Enterprise". At the time, we just needed a name quickly, but it was never a good name, just a placeholder. This removes that placeholder name from much of the codebase. A bit more work will be required to transition the `zulip::voyager` Puppet class, as that has some migration work involved.	2020-01-30 12:40:41 -08:00
Tim Abbott	d70e799466	bots: Remove FEEDBACK_BOT implementation. This legacy cross-realm bot hasn't been used in several years, as far as I know. If we wanted to re-introduce it, I'd want to implement it as an embedded bot using those common APIs, rather than the totally custom hacky code used for it that involves unnecessary queue workers and similar details. Fixes #13533.	2020-01-25 22:41:39 -08:00
Mateusz Mandera	8acfa17fe6	models: Add recipient foreign key in UserProfile and Stream. This is adds foreign keys to the corresponding Recipient object in the UserProfile on Stream tables, a denormalization intended to improve performance as this is a common query. In the migration for setting the field correctly for existing users, we do a direct SQL query (because Django 1.11 doesn't provide any good method for doing it properly in bulk using the ORM.). A consequence of this change to the model is that a bit of code needs to be added to the functions responsible for creating new users (to set the field after the Recipient object gets created). Fortunately, there's only a few code paths for doing that. Also an adjustment is needed in the import system - this introduces a circular relation between Recipient and UserProfile. The field cannot be set until the Recipient objects have been created, but UserProfiles need to be created before their corresponding Recipients. We deal with this by first importing UserProfiles same way as before, but we leave the personal_recipient field uninitialized. After creating the Recipient objects, we call a function to set the field for all the imported users in bulk. A similar change is made for managing Stream objects.	2019-12-09 15:14:41 -08:00
Mateusz Mandera	dbe508bb91	models: Migration of Message.pub_date to date_sent, part 2. Fixes #1727. With the server down, apply migrations 0245 and 0246. 0246 will remove the pub_date column, so it's essential that the previous migrations ran correctly to copy data before running this.	2019-10-05 19:01:34 -07:00
Tim Abbott	02d55928ea	import: Fix importing slack avatars into S3_UPLOAD_BACKEND. Apparently, a subtle mismatch between the filename/URL formats for our upload codebases meant that importing Slack avatars into systems using S3_UPLOAD_BACKEND would end up with the avatars having the wrong URLs.	2019-07-21 21:25:31 -07:00
Tim Abbott	fd25ced43c	import: Fix check for whether data-user-id is present. Apparently, the `in` keyword in Beautiful Soup does something different.	2019-06-18 11:13:32 -07:00
Tim Abbott	649e363ee3	import: Fix handling of legacy and wildcard mentions. Our recently-added code for rewriting user IDs on data import didn't correctly handle wildcard mentions and mentions generated by very old versions of Zulip (pre data-user-id).	2019-06-18 10:35:01 -07:00
Tim Abbott	2538f84447	import: Fix bad database query for first_message_id. The previous query ended up doing an awkward join that did not guarantee use of the Recipient index on zerver_message, turning a very fast query into something that could take much longer for a single stream than the rest of the import combined.	2019-06-18 10:25:00 -07:00
Vishnu Ks	8718846c2a	import: Use html.parser instead of lxml in bs4. lxml parser appends html and body tags to the soup object which are not reqired. There are no other major parsing diffrences between the two parsers as long the HTML input is perfectly formated. lxml parser is much faster than html.parser but it hardly matters in our case. https://www.crummy.com/software/BeautifulSoup/bs4/doc/#differences- between-parsers	2019-06-02 14:53:13 -07:00
Vishnu Ks	31151dadbf	import: Replace data-user-group-id in rendered_content. See the data-user-id commit for details.	2019-05-28 12:53:20 -07:00
Vishnu Ks	ce1d6044db	import: Replace data-stream-id in rendered_content. See the data-user-id commit for details.	2019-05-28 12:53:20 -07:00
Vishnu Ks	cb5b3f347b	import: Replace data-user-id in rendered_content with new user id. Previously, if you exported a Zulip organization and then re-imported it, we'd end up renumbering the user IDs and all direct foreign key references to them in the database, but not the data-user-id references in mentions. Fix this by parsing the message content and doing that renumbering. (Because we import raw markdown, not HTML, from third-party tools, these changes won't affect data import from slack etc.) Fixes the high-priority part of #11293.	2019-05-28 12:53:19 -07:00
Anders Kaseorg	643bd18b9f	lint: Fix code that evaded our lint checks for string % non-tuple. Signed-off-by: Anders Kaseorg <anders@zulipchat.com>	2019-04-23 15:21:37 -07:00
Challa Venkata Raghava Reddy	b69aec2dbc	streams: Add first_message_id tracking first message in stream. This field is primarily intended to support avoiding displaying the "more topics" feature in new organizations and streams, where we might know that all messages in the stream are already available in the browser. Based on original work by Roman Godov, and significantly modified by tabbott. The second migration involved here could be expensive on Zulip Cloud, but is unlikely to be an issue on other servers.	2019-03-11 13:30:49 -07:00
Hemanth V. Alluri	ae126c452b	stream-descriptions: Create wrapper for rendering stream descriptions. In commit `de65a04` we can see that if the need ever arises to modify how stream descriptions are rendered, we would need to make changes at 5 different call points which can be quite cumbersome. So this functionality has been extracted to a new method called 'render_stream_descriptions'.	2019-03-06 17:16:14 -08:00
Bennet Sunder	7c5f316cb8	alert_words: Performance improvements in looking for alert_words. This commit leverages the ahocorasick algorithm to build a set of user_ids that have their alert_words present in the message. It runs in linear time of the order of length of the input message as opposed to number of alert_words. This is after building a ahocorasick Automaton which runs in O(number of alert_words in entire realm) which is usually cached.	2019-03-01 15:36:39 -08:00
Tim Abbott	de65a04ae0	streams: Disable inline URL preview when rendering stream descriptions. We want to use the baseline features of bugdown, but not fancy things like inline URL previews, since the whole structure of stream descriptions is to have a single-line thing supporting some formatting. The migration part of this change fixes a bug encountered by some organizations upgrading from older versions of Zulip.	2019-02-28 17:00:40 -08:00
Tim Abbott	4d08461ab1	import: Set plan_type to SELF_HOSTED on import. We've for a while had logic to set plan_type to LIMITED when importing into Zulip Cloud; we need corresponding logic to set it to SELF_HOSTED when importing into a self-hosted server. Fixes #11541.	2019-02-12 16:01:02 -08:00
Wyatt Hoodes	9c68a97472	import/export: Use separate analytics.json for analytics data. This helps keep the realm.json small and easy to process; previously, almost the entire size of that file was the analytics data. We implement this by refactoring the analytics Config objects into a separate subroutine that writes to a separate file, plus the corresponding import code. Manual testing was performed by exporting the 'analytics' realm, and importing back to a newly created 'test' realm. The 'test' realm was then exported and the json files were inspected. The data appeared consistent with no abnormalities. Fixes: #11220.	2019-02-04 10:59:24 -08:00
Hemanth V. Alluri	73d26c8b28	streams: Render and store the stream description from the backend. This commit does the following three things: 1. Update stream model to accomodate rendered description. 2. Render and save the stream rendered description on update. 3. Render and save stream descriptions on creation. Further, the stream's rendered description is also sent whenever the stream's description is being sent. This is preparatory work for eliminating the use of the non-authoritative marked.js markdown parser for stream descriptions.	2019-02-01 22:24:18 -08:00
Vishnu Ks	bec875a9af	import realm: Use processes for resizing avatar images. This should significantly improve the data import performance when importing large open source realms from Slack. Fixes #11009.	2019-01-25 12:37:12 -08:00
Tim Abbott	dfaa2e481d	import: Log a warning when avatars can't be thumbnailed. This fixes a potential crash in the import tool if a single user has a broken avatar image.	2019-01-15 16:48:04 -08:00
Tim Abbott	6eda129741	export: Export and import analytics table data. This should eliminate the need to do manual analytics work when importing organizations imported/exported using the zulip -> zulip import/export tools.	2019-01-04 16:22:18 -08:00
Tim Abbott	48ccb3ad18	import: Move realm_tables to the appropriate file. These had ended up in the wrong place when we split export from import.	2019-01-04 16:22:18 -08:00
Tim Abbott	b33e0ad539	import: Fix pointer logic to sort by message_id. Previously, the pointer calculation logic wasn't sorting by message ID, which caused the database queries to not properly use the indexes they should.	2019-01-04 16:22:18 -08:00
Tim Abbott	a1919971e4	import: Handle invalid data-user-id values for mentions. This is an issue with zulip -> zulip server data imports.	2019-01-02 15:23:09 -08:00
Tim Abbott	b63f8b59b2	import: Handle corner case around EMAIL_GATEWAY_BOT emails.	2019-01-02 15:23:09 -08:00
Tim Abbott	8cfea958de	import: Fix pointer logic for zulip->zulip imports. Previously, the pointer was almost guaranteed to be an invalid random value, because we renumber message IDs unconditionally now.	2019-01-02 15:23:09 -08:00
Tim Abbott	74ff77d366	import: Always set a valid content-type for S3 backend. The octet-stream content type is potentially under-specified, but it's better than potentially submitting None and increases consistency of this part of the codebase.	2018-12-29 22:13:11 -08:00
Tim Abbott	f0c7424957	import: Fix sending floats to boto S3 metadata keys. The boto library's s3 interface allows setting only string-format metadata keys. So we need to cast the last_modified floating-point timestamp into a string before storing on the S3 object. This bug mostly broke uploading avatars when using the S3 storage backend.	2018-12-29 22:09:31 -08:00
Tim Abbott	c995e8e2ae	import: Ensure presence of basic avatar images for HipChat. Our HipChat conversion tool didn't properly handle basic avatar images, resulting in only the medium-size avatar images being imported properly. This fixes that bug by asking the import tool to do the thumbnailing for the basic avatar image (from the .original file) as well as the medium avatar image.	2018-12-27 17:47:09 -08:00
Rishi Gupta	8a95526ced	billing: Always transition to Realm.LIMITED via do_change_plan_type. Fixes a bug in import_realm where secondary attributes like message visibility weren't being set, and also makes bugs like this less likely in the future. Also, putting the plan_type change at the end of import_realm, so that future restrictions to LIMITED realms don't affect the import process.	2018-12-13 13:26:24 -08:00
Tim Abbott	1adc40f014	import: Deduplicate functions for uploading to S3/files. We've had a long stream of bugs existed because only one of these two code paths was tested (usually the local uploads backend). By deduplicating these functions, we ensure that this category of bugs no longer happens. Following my recent refactor, this is just a straightforward merge, with code for one or the other backend ending up inside an if statement.	2018-12-05 16:15:01 -08:00
Tim Abbott	c9b801efde	import: Use the s3_path attribute for path_maps unconditionally. While the s3_path is almost always the same as the path, structurally, `path` is the location in the export object, whereas s3_path is the URL path.	2018-12-05 16:15:01 -08:00
Tim Abbott	f4c5a45f4f	import: Fix S3 paths for imported avatar PNG. Previously, we were incorrectly importing avatar PNGs to a filename without the .png extension, resulting in them effectively not being imported. This was mitigated by the fact that we imported the originals and ran the appropriate `ensure_` functions, but still a bug.	2018-12-05 16:15:01 -08:00
Tim Abbott	412dc8dcda	import: Set last_modified in import_uploads_local. This has no effect other than to make the S3 and local code paths more nearly identical.	2018-12-05 16:15:01 -08:00
Tim Abbott	d8d0492d64	import: Restructure uploads path logic to be more similar. This is preparation for future deduplication of the two redundant uploads backends.	2018-12-05 16:15:01 -08:00
Tim Abbott	671ceccd78	import: Deduplicate medium avatars special logic. This requires a bit of care with upload_backend to avoid breaking how we mock that class in our tests.	2018-12-05 16:15:01 -08:00
Tim Abbott	36b43a6d7a	import: Deduplicate first block of import_uploads logic.	2018-12-05 16:15:01 -08:00
Tim Abbott	f80bab58c0	import_realm: Add progress indicator for importing uploads. This makes it easier to see how we're doing when uploading a very large number of files.	2018-12-05 16:15:01 -08:00
Steve Howell	88f50b97fd	import: Render content before inserting messages. By rendering content before bulk importing messages, we avoid O(N) database hops.	2018-11-07 10:33:11 -08:00
Steve Howell	bf3f7d93d0	Simplify params for fix_message_rendered_content.	2018-11-07 10:33:11 -08:00
Steve Howell	0878d86706	import: Avoid unnecessary Message lookups. We now no longer go the DB to get a Message object during render.	2018-11-07 10:33:11 -08:00
Steve Howell	1e12b13a56	import: Avoid unnecessary sender lookups. This commit speeds up the import by avoiding sender lookups and instead using the data for users that we already have in memory. This avoids a few DB hops, many hops to memcached, plus some object construction. We now call do_render_markdown() directly. This also makes it more explicit that the import has never rendered alert words.	2018-11-07 10:33:10 -08:00
Steve Howell	f9a7451167	import: Pass in realm to render codepath. We avoid querying the same realm multiple times.	2018-11-07 10:08:46 -08:00

1 2

96 Commits