notmuch - thread-based email index, search and tagging

	Commit message (Collapse)	Author	Age
...
*	lib: Simplify code flow in _resolve_message_id_to_thread_id	Carl Worth	2010-04-12
\| \| \| \| \| \|	There are two primary cases in this function, (the message exists in the database or it does not). Previously the code for these two cases was split and intermingled with goto-spaghetti connections.
*	lib: Fix internal documentation of _resolve_message_id_to_thread_id	Carl Worth	2010-04-12
\| \| \| \| \|	We no longer return NULL, but instead generate a new thread ID for messages that we haven't seen yet.
*	Store thread ids for messages that we haven't seen yet	James Westby	2010-04-12
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	This allows us to thread messages even when we receive them out of order, or never receive the root. The thread ids for messages that aren't present but are referred to are stored as metadata in the database and then retrieved if we ever get that message. When determining the thread id for a message we also check for this metadata so that we can thread descendants of a message together before we receive it. Edited by Carl Worth <cworth@cworth.org>: Split this portion of the commit from the earlier-applied portion adding test cases.
*	Add is:<tag> as a synonym for tag:<tag> in search terms.	Carl Worth	2010-03-09
\| \| \| \| \|	I like the readability of this, it provides compatibility with people trained in this syntax by sup, and it even saves one character.
*	lib: Rename iterator functions to prepare for reverse iteration.	Carl Worth	2010-03-09
\| \| \| \| \| \| \| \|	We rename 'has_more' to 'valid' so that it can function whether iterating in a forward or reverse direction. We also rename 'advance' to 'move_to_next' to setup parallel naming with the proposed functions 'move_to_first', 'move_to_last', and 'move_to_previous'.
*	Fix printf for when uint64_t != unsigned long long int	Carl Worth	2010-02-09
\| \| \| \| \| \| \| \|	Thanks to Michal Sojka <sojkam1@fel.cvut.cz> for pointing out the correct fix, which I verified in the freely-available WG14/N1124 draft (from the C99 working group) which is available here: http://www.open-std.org/JTC1/SC22/wg14/www/docs/n1124.pdf
*	Switch from random to sequential thread identifiers.	Carl Worth	2010-02-09
\| \| \| \| \| \| \| \| \| \| \| \| \|	The sequential identifiers have the advantage of being guaranteed to be unique (until we overflow a 64-bit unsigned integer), and also take up half as much space in the "notmuch search" output (16 columns rather than 32). This change also has the side effect of fixing a bug where notmuch could block on /dev/random at startup (waiting for some entropy to appear). This bug was hit hard by the test suite, (which could easily exhaust the available entropy on common systems---resulting in large delays of the test suite).
*	notmuch new: Print upgrade progress report as a percentage.	Carl Worth	2010-01-09
\| \| \| \| \| \| \| \| \| \| \| \|	Previously we were printing a number of messages upgraded so far. The original motivation for this was to accurately reflect the fact that there are two passes, (so each message is processed twice and it's not accurate to represent with a single count). But as it turns out, the second pass takes zero time (relatively speaking) so we're still not accounting for it. If nothing else, the percentage-based reporting makes for a cleaner API for the progress_notify function.
*	lib: Split the database upgrade into two phases for safer operation.	Carl Worth	2010-01-09
\| \| \| \| \| \| \| \| \|	The first phase copies data from the old format to the new format without deleting anything. This allows an old notmuch to still use the database if the upgrade process gets interrupted. The second phase performs the deletion (after updating the database version number). If the second phase is interrupted, there will be some unused data in the database, but it shouldn't cause any actual harm.
*	lib: Delete stale timestamp documents during database upgrade.	Carl Worth	2010-01-08
\| \| \| \| \|	Once we move the timestamp to the new directory document, we don't need the old one anymore.
*	notmuch new: Fix progress notification on database upgrade.	Carl Worth	2010-01-07
\| \| \| \| \|	This was firing continuously rather than just once per second as intended.
*	lib: Implement versioning in the database and provide upgrade function.	Carl Worth	2010-01-07
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	The recent support for renames in the database is our first time (since notmuch has had more than a single user) that we have a database format change. To support smooth upgrades we now encode a database format version number in the Xapian metadata. Going forward notmuch will emit a warning if used to read from a database with a newer version than it natively supports, and will refuse to write to a database with a newer version. The library also provides functions to query the database format version: notmuch_database_get_version to ask if notmuch wants a newer version than that: notmuch_database_needs_upgrade and a function to actually perform that upgrade: notmuch_database_upgrade
*	Prefer READ_ONLY consistently over READONLY.	Carl Worth	2010-01-07
\| \| \| \| \| \|	Previously we had NOTMUCH_DATABASE_MODE_READ_ONLY but NOTMUCH_STATUS_READONLY_DATABASE which was ugly and confusing. Rename the latter to NOTMUCH_STATUS_READ_ONLY_DATABASE for consistency.
*	lib: Consolidate checks for read-only database.	Carl Worth	2010-01-07
\| \| \| \| \| \| \| \| \| \| \|	Previously, many checks were deep in the library just before a cast operation. These have now been replaced with internal errors and new checks have instead been added at the beginning of all top-levelentry points requiring a read-write database. The new checks now also use a single function for checking and printing the error message. This will give us a convenient location to extend the check, (such as based on database version as well).
*	lib: Clarify internal documentation of _notmuch_database_filename_to_direntry	Carl Worth	2010-01-07
\| \| \| \| \| \| \|	The original wording made it sound like this function was just doing some string manipulation. But this function actually creates new directory documents as a side effect. So make that explicit in its documentation.
*	lib: Indicate whether notmuch_database_remove_message removed anything.	Carl Worth	2010-01-06
\| \| \| \| \| \| \| \|	Similar to the return value of notmuch_database_add_message, we now enhance the return value of notmuch_database_remove_message to indicate whether the message document was entirely removed (SUCCESS) or whether only this filename was removed and the document exists under other filenamed (DUPLICATE_MESSAGE_ID).
*	Add missing comment for NOTMUCH_STATUS_READONLY_DATABASE.	Carl Worth	2010-01-06
\| \| \| \|	And adjust the string representation of the same to match.
*	lib: Implement new notmuch_directory_t API.	Carl Worth	2010-01-06
\| \| \| \| \| \| \|	This new directory ojbect provides all the infrastructure needed to detect when files or directories are deleted or renamed. There's still code needed on top of this (within "notmuch new") to actually do that detection.
*	database: Add new, public notmuch_database_remove_message	Carl Worth	2010-01-06
\| \| \| \| \| \|	This will allow applications to support the removal of messages, (such as when a file is deleted from the mail store). No removal support is provided yet in commands such as "notmuch new".
*	database: Add new find_doc_ids_for_term interface.	Carl Worth	2010-01-06
\| \| \| \| \| \| \| \|	The existing find_doc_ids function is convenient when the caller doesn't want to be bothered constructing a term. But when the caller does have the term already, that interface is just wasteful. So we export a lower-level interface that maps a pre-constructed term to a document-ID iterators.
*	database: Make find_unique_doc_id enforce uniqueness (for a debug build)	Carl Worth	2010-01-06
\| \| \| \| \|	Catching any violation of this unique-ness constraint is very much in line with similar, existing INTERNAL_ERROR cases.
*	database: Abstract _filename_to_direntry from _add_message	Carl Worth	2010-01-06
\| \| \| \| \| \|	The code to map a filename to a direntry is something that we're going to want in a future _remove_message function, so put it in a new function _notmuch_database_filename_to_direntry .
*	database: Allowing storing multiple filenames for a single message ID.	Carl Worth	2010-01-06
\| \| \| \| \| \|	The library interface is unchanged so far, (still just notmuch_database_add_message), but internally, the old _set_filename function is now _add_filename instead.
*	database: Store mail filename as a new 'direntry' term, not as 'data'.	Carl Worth	2010-01-06
\| \| \| \| \| \| \| \| \| \|	Instead of storing the complete message filename in the data portion of a mail document we now store a 'direntry' term that contains the document ID of a directory document and also the basename of the message filename within that directory. This will allow us to easily store multple filenames for a single message, and will also allow us to find mail documents for files that previously existed in a directory but that have since been deleted.
*	database: Split _find_parent_id into _split_path and _find_directory_id	Carl Worth	2010-01-06
\| \| \| \| \| \|	Some pending commits want the _split_path functionality separate from mapping a directory to a document ID. The split_path function now returns the basename as well as the directory name.
*	database: Store directory path in 'data' of directory documents.	Carl Worth	2010-01-06
\| \| \| \| \| \| \| \|	We're planning to have mail documents refer to directory documents for the path of the containing directory. To support this, we need the path in the data, (since the path in the 'directory' term can be irretrievable as it will be the SHA1 sum of the path for a very long path).
*	database: Export _notmuch_database_find_parent_id for internal use.	Carl Worth	2010-01-06
\| \| \| \| \| \|	We'll soon have mail documents referring to their parent directory's directory documents, so we'll need access to _find_parent_id in files such as message.cc.
*	database: Store the parent ID for each directory document.	Carl Worth	2010-01-06
\| \| \| \| \| \| \|	Storing the document ID of the parent of each directory document will allow us to find all child-directory documents for a given directory document. We will need this in order to detect directories that have been removed from the mail store, (though we aren't yet doing this).
*	database: Rename internal directory value from XTIMESTAMP to XDIRECTORY.	Carl Worth	2010-01-06
\| \| \| \| \| \| \| \|	The recent change from storing absolute paths to relative paths means that new directory documents will already be created, (and the old ones will just linger stale in the database). Given that, we might as well put a clean name on the term in the new documents, (and no real flag day is needed).
*	database: Store directory paths as relative, not absolute.	Carl Worth	2010-01-06
\| \| \| \| \| \| \|	We were already storing relative mail filenames, so this is consistent with that. Additionally, it means that directory documents remain valid even if the database is relocated within its containing filesystem.
*	lib: Document that the filename is stored in the 'data' of a mail document	Carl Worth	2010-01-06
\| \| \| \| \| \|	Our database schema documentation previously didn't give any indication of where this most essential piece of information is stored.
*	lib: Rename set/get_timestamp to set/get_directory_mtime.	Carl Worth	2010-01-06
\| \| \| \| \| \|	I've been suitably scolded by Keith for doing a premature generalization that ended up just making the documentation more convoluted. Fix that.
*	lib: Abstract the extraction of a relative path from set_filename	Carl Worth	2010-01-06
\| \| \| \| \| \|	We'll soon be having multiple entry points that accept a filename path, so we want common code for getting a relative path from a potentially absolute path.
*	Nuke the remainings of _notmuch_message_add_thread_id.	Fernando Carrijo	2009-12-09
\| \| \| \| \| \| \| \| \| \|	The function _notmuch_message_add_thread_id has been removed from the private interface of notmuch. There's no reason for one to keep a declaration of its prototype in the code base. Also, lets update a commentary that referenced that function and escaped from previous scrutiny. Signed-off-by: Fernando Carrijo <fcarrijo@yahoo.com.br>
*	notmuch: New function to retrieve all tags from the database.	Jan Janak	2009-11-26
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	This patch adds a new function called notmuch_database_get_all_tags which can be used to obtain a list of all tags from the database (in other words, the list contains all tags from all messages). The function produces an alphabetically sorted list. To add support for the new function, we rip the guts off of notmuch_message_get_tags and put them in a new generic function called _notmuch_convert_tags. The generic function takes a Xapian::TermIterator as argument and uses the iterator to find tags. This makes the function usable with different Xapian objects. Function notmuch_message_get_tags is then reimplemented to call the generic function with message->doc.termlist_begin() as argument. Similarly, we implement notmuch_message_database_get_all_tags, the function calls the generic function with db->xapian_db->allterms_begin() as argument. Finally, notmuch_database_get_all_tags is exported through lib/notmuch.h Signed-off-by: Jan Janak <jan@ryngle.com>
*	Add rudimentary date-based search.	Carl Worth	2009-11-23
\| \| \| \| \| \| \| \| \| \| \| \| \|	The rudimentary aspect here is that the date ranges are specified with UNIX timestamp values (number of seconds since 1970-01-01 UTC). One thing that can help here is using the date program to determins timestamps, such as: $(date +%s -d 2009-10-01)..$(date +%s) Long-term, we'll probably need to do our own query parsing to be able to support directly-specified dates and also relative expressions like "since:'2 months ago'".
*	lib/database.cc: coding style	Chris Wilson	2009-11-22
\| \| \| \| \| \|	Carl claims he must have been distracted when he wrote this... Signed-off-by: Chris Wilson <chris@chris-wilson.co.uk>
*	add_message: Use sha-1 in place of overly long message ID.	Carl Worth	2009-11-22
\| \| \| \| \| \| \| \| \| \| \| \| \|	Since Xapian has a limit on the maximum length of a term, we have to check for that before trying to add the message ID as a term. This fixes the bug reported by Mike Hommey here: <20091120132625.GA19246@glandium.org> I've also constructed 20 files with a range of message ID lengths centered around the Xapian term-length limit which I'll use to seed a new test suite soon.
*	get_timestamp: Ensure that return value is 0 in case of exception.	Carl Worth	2009-11-22
\| \| \| \|	Just to be on the safe side of things.
*	Catch and optionally print about exception at database->flush.	Carl Worth	2009-11-22
\| \| \| \| \| \| \|	If an earlier exception occurred, then it's not unexpected for the flush to fail as well. So in that case, we'll silently catch the exception. Otherwise, make some noise about things going wrong at the time of flush.
*	Print information about where Xapian exception occurred.	Carl Worth	2009-11-22
\| \| \| \| \|	Previously, our Xapian exception reports where identical so they were hard to track down.
*	Rename NOTMUCH_DATABASE_MODE_WRITABLE to NOTMUCH_DATABASE_MODE_READ_WRITE	Carl Worth	2009-11-21
\| \| \| \|	And correspondingly, READONLY to READ_ONLY.
*	Permit opening the notmuch database in read-only mode.	Chris Wilson	2009-11-21
\| \| \| \| \| \| \| \| \|	We only rarely need to actually open the database for writing, but we always create a Xapian::WritableDatabase. This has the effect of preventing searches and like whilst updating the index. Signed-off-by: Chris Wilson <chris@chris-wilson.co.uk> Acked-by: Carl Worth <cworth@cworth.org>
*	add_message: Re-fix handling of non-mail files.	Carl Worth	2009-11-20
\| \| \| \| \| \| \| \| \| \| \|	More fallout from _get_header now returning "" for missing headers. The bug here is that we would no longer detect that a file is not an email message and give up on it like we should. And this time, I actually audited all callers to notmuch_message_get_header, so hopefully we're done fixing this bug over and over.
*	notmuch_database_add_message: Add missing error-value propagation.	Carl Worth	2009-11-20
\| \| \| \| \|	Thanks to Mike Hommey for doing the analysis that led to noticing that this was missing.
*	add_message: Properly handle missing Message-ID once again.	Carl Worth	2009-11-20
\| \| \| \| \| \| \| \|	There's been a fair amount of fallout from when we changed message_file_get_header from returning NULL to returning "" for missing headers. This is yet more fallout from that, (where we were accepting an empty message-ID rather than generating one like we want to).
*	Typsos	Ingmar Vanhassel	2009-11-18
\|
*	linke_message: Avoid segfault when In-Reply-to header is empty.	Carl Worth	2009-11-18
\| \| \| \| \| \| \| \| \| \| \| \| \|	This was recently introduced in commit: 64c03ae97f2f5294c60ef25d7f41849864e6ebd3 which was adding extra checks to avoid adding a self-referencing message. How many times am I going to fix a dumb regression like this and say "we really need a test suite" before I actually sit down and write the test suite?
*	database: Make _parse_message_id static once again.	Carl Worth	2009-11-17
\| \| \| \| \| \| \|	We had exposed this to the internal implementation for a short time, (only while we had the silly code fetching In-Reply-To values from message files instead of from the database). Make this private again as it should be.
*	database: Add "replyto" to the database schema documentation.	Carl Worth	2009-11-17
\| \| \| \| \| \|	Maybe ths lack of this documentation is why I forgot we were actually storing this and wrote the ugly code to fetch In-Reply-To from message files rather than from the database.