zulip

Commit Graph

Author	SHA1	Message	Date
Rishi Gupta	e183c316dd	help: Rename help/change-your-avatar to help/set-your-avatar.	2019-02-13 17:50:39 -08:00
Anders Kaseorg	56a675d5ec	export: Remove unused imports. Signed-off-by: Anders Kaseorg <andersk@mit.edu>	2019-02-02 17:25:27 -08:00
Pragati Agrawal	e1772b3b8f	tools: Upgrade Pycodestyle and fix new linter errors. Here, we are upgrading pycodestyle version from 2.4.0 to 2.5.0. Fixes: #11396.	2019-01-31 12:21:41 -08:00
Matthew Wegner	370cf1a2cb	import: Normalize Slackbot String Comparison. In very old Slack workspaces, slackbot can appear as "Slackbot", and the import script only checks for "slackbot" (case sensitive). This breaks the import--it throws the assert that immediately follows the test. I don't know how common this is, but it definitely affected our import. The simple fix is to compare against a lowercased-version of the user's full name.	2019-01-28 14:59:41 -08:00
Tim Abbott	8a90441d2f	slack import: Import long-inactive users as long-term idle. This avoids creating UserMessage rows for long-inactive users in organizations with many thousands of users.	2018-12-16 18:52:20 -08:00
Tim Abbott	a6ca95dfc4	slack import: Fix all messages being imported to one channel. This was an ugly variable-escape-from-loop regression introduced in `e59ff6e6db`.	2018-12-12 17:54:37 -08:00
Tim Abbott	d6217eb862	slack import: Fix empty values for custom profile fields. The Slack import process would incorrectly issue CustomProfileFieldValue entries with a value of "" for users who didn't have a given CustomProfileField (especially common for the "skype" and "phone" fields). This had no user-visible effect, but certainly added some clutter in the database.	2018-12-12 12:58:27 -08:00
rht	e59ff6e6db	slack import: Eliminate need to load all messages into memory. This works by yielding messages sorted based on timestamp. Because the Slack exports are broken into files by date, it's convenient to do a 2-layer sorting process, where we open all the files for a given day, and then sort their messages by timestamp before yielding them. Fixes #10930.	2018-12-05 12:20:50 -08:00
Steve Howell	d86dd165da	gitter/slack/hipchat: Remove "subject" from conversions. We (lexically) remove "subject" from the conversion code. The `build_message` helper calls `set_topic_name` under the hood, so things still have "subject" in the JSON. There was good code coverage on `build_message`.	2018-11-12 15:47:11 -08:00
Tim Abbott	8b661f2f03	slack import: Correctly detect the commenting user. Fixes #10772.	2018-11-06 13:14:23 -08:00
Steve Howell	30c493ed24	slack import: Generate message_id/reaction_id with NEXT_ID. This avoids the need to pass tuples of ints around, which is pretty brittle.	2018-10-29 13:24:50 -07:00
Steve Howell	2f58eb1057	slack import: Extract process_message_files(). This is mostly an extraction, but it does change the way we calculate `content`. We append the markdown links from ALL files to any content that came in the message itself. Separating this out also allows us to add more test coverage for the extracted code.	2018-10-29 13:24:50 -07:00
Steve Howell	00f822a26a	conversion: Generate attachment_ids with helpers.	2018-10-29 13:24:50 -07:00
Steve Howell	5cb60f7bea	conversions: Use subscriber_map for Slack/Gitter. We now use subscriber_map for building UserMessage rows in Slack/Gitter conversions. This is mostly designed to simplify the code, rather than having to scan the entire subscribers for each message. I am guessing this will improve performance for most conversions. We sort small lists on every message, in order to be deterministic, but the sorting cost is probably more than offset by avoiding the O(N) scans across all subscriptions. Also, it's probably negligible in the grand scheme of things, compared to JSON parsing, file I/O, etc. This commits also fixes some typos with mentioned_users_id -> mentioned_user_ids and cleans up a test a bit as well.	2018-10-29 13:24:50 -07:00
Steve Howell	5194701787	conversions: Use NEXT_ID for usermessage_id. This is mostly complicated due to the way that the Slack import passes around tuples of ids to maintain four different parallel sequences.	2018-10-29 13:24:50 -07:00
Tim Abbott	f9b6eeb488	import: Migrate from json to ujson for better perf. We expect to get better memory performace from ujson than json. We also do a better job of closing file handles. This likely fixes #10377.	2018-10-17 12:11:08 -07:00
Tim Abbott	78a15dd715	slack import: Fix obscure email address for Slackbot. Since we know what slackbot is, we don't need to give it a crazy hash as its email address.	2018-10-16 16:33:41 -07:00
Steve Howell	8accc60ca7	import_util: Support multiple message ids for attachments.	2018-10-13 16:47:44 -07:00
Steve Howell	23d7b3d2cc	import: De-dup create_converted_data_files helper.	2018-10-13 16:47:41 -07:00
Rhea Parekh	f70b9a3eba	import: Move 'build_message' to import_util.	2018-08-19 22:27:13 -07:00
Rhea Parekh	53e9da8e1f	import: Build CustomProfileField, CustomProfileFieldValue and RealmEmoji with model class.	2018-08-19 22:27:13 -07:00
Rhea Parekh	d98a5925cb	import: Build Reaction with the model class.	2018-08-19 22:27:13 -07:00
Rhea Parekh	a5bc701181	import: Move 'build_stream' to import_util.	2018-08-19 22:27:13 -07:00
Rhea Parekh	c4f8abbd30	import: Build Message with the model class.	2018-08-19 22:27:13 -07:00
Rhea Parekh	4ea7302e14	import: Add missing fields in UserProfile object. The missing fields are checked by `full_clean()` method. The datetime field errors are ignored as they are fixed in the `import_realm` script. The field that are allowed to be null are not included while building this object.	2018-08-19 22:27:13 -07:00
Rhea Parekh	c77763bd8e	import: Move 'build_realm' to import_util.	2018-08-19 22:27:13 -07:00
Tim Abbott	8a22838acf	slack import: Fix computation of owner email for uploaded files. The previous code was just always returning the first user in the organization, due to an incorrect comparison.	2018-08-10 16:20:36 -07:00
Rhea Parekh	3ff339c294	slack import: Add support for uploads in messages through 'files' keyword. It appears that Slack just changed their export format, and how uses this `files` list for user-uploaded files.	2018-08-10 16:20:36 -07:00
Rhea Parekh	20bca1409f	import: Set emoji records 'last_modified' value in 'import_uploads_s3'. The 'last_modified' value in emoji records is needed for uploading the file to the S3 backend. We set the same in the function 'import_uploads_s3'. We also have to remove the keyword 'last_modified' while building the RealmEmoji dict, as it is not a field which exists in RealmEmoji objects.	2018-08-10 16:20:36 -07:00
Tim Abbott	cf8a0ae819	slack import: Set a last_modified timestamp for custom emoji.	2018-08-10 09:27:43 -07:00
Rhea Parekh	18a4904437	import: Move 'build_attachment' to import_util.	2018-08-07 16:45:42 -07:00
Rhea Parekh	b6ccc0bc52	import: Move 'build_defaultstream' to import_util.	2018-08-07 16:45:42 -07:00
Rhea Parekh	bee3964f14	import: Move 'build_usermessages' to import_util.	2018-08-07 16:45:42 -07:00
Rhea Parekh	eefe7cccd2	import: Move 'process_uploads' and 'process_emojis' to import_util.	2018-08-07 16:45:42 -07:00
Rhea Parekh	30cc7354eb	import: Move 'process_avatars' to import_util.	2018-08-07 16:45:40 -07:00
Rhea Parekh	87cc1a6280	import: Move 'build_subscription' and 'build_recipient' to import_util.	2018-08-07 16:35:56 -07:00
Rhea Parekh	a516f80646	import: Move 'build_avatar' to import_util.	2018-08-07 16:35:56 -07:00
Rhea Parekh	1117455a90	import: Move 'ZerverFieldsT' and 'build_zerver_realm' to import_util.	2018-08-07 16:35:56 -07:00
Rhea Parekh	b8e1e8b31d	import: Add slack import files in zerver/data_import directory.	2018-08-01 11:52:14 -07:00

39 Commits