mirrors/jj

mirror of https://github.com/martinvonz/jj.git synced 2025-01-16 09:11:55 +00:00

Author	SHA1	Message	Date
Martin von Zweigbergk	70b99c960e	transaction: make commit() return resulting ReadonlyRepo I've wanted the API to look like this for a while. It seems like a good API to me. It means that the caller won't have to reload the repo after committing. The cost seems relatively small. It involves copying potentially a lot of data in memory (at least the View object), but it shouldn't involve reading from disk or any other processing. To reduce the amount of data to copy, it may be worth switching to persistent data types. I've also wanted to do that for the copying we do when start a transaction. I couldn't measure any slowdown caused by this change.	2021-05-08 13:50:59 -07:00
Martin von Zweigbergk	9e3e6f03a1	repo: store Operation object, not just its ID in ReadonlyRepo This change simplifies a bit on its own, and it will help with the next change as well.	2021-05-07 22:53:11 -07:00
Martin von Zweigbergk	c05878b9a6	maintenance: add support for latest nightly toolchain	2021-05-07 22:48:24 -07:00
Martin von Zweigbergk	eb348be32c	revsets: make description() lazy by using new FilterRevset	2021-05-02 15:16:51 -07:00
Martin von Zweigbergk	e6f751a24d	revsets: generalize merge revset implementation to be a generic filter	2021-05-02 14:58:10 -07:00
Martin von Zweigbergk	7065cecfdc	revsets: add revset yielding merge commits	2021-05-02 14:33:38 -07:00
Martin von Zweigbergk	aef27d5701	revsets: remove transitive edges in graph iterator by default The git.git repo seems to have lots of merges from far back in the history into newer history. That results in `jj log -r 'git_refs()'` being completely useless because of the number of such edges. For example, v2.31.0 has almost 600 edges going out of it and presumably merging (forking) back into various different previous versions. Git, unlike Mercurial, seems to remove an edge from the graph if the edge can also be reached via a longer path. This commit makes it so we also do that (i.e. the filtered graph is a transitive reduction of the graph before filtering). This slows down `jj log -r ,,v2.0.0 -T ""` by about 2%. That's still small enough that it doesn't seem worth it to have a separate iterator for contiguous ranges (which would be an option).	2021-05-01 23:25:33 -07:00
Martin von Zweigbergk	33da97f0bf	revsets: add iterator adapter for rendering simplified graph of set When rendering a non-contiguous subset of the commits, we want to still show the connections between the commits in the graph, even though they're not directly connected. This commit introduces an adaptor for the revset iterators that also yield the edges to show in such a simplified graph. This has no measurable impact on `jj log -r ,,v2.0.0` in the git.git repo. The output of `jj log -r 'v1.0.0 \| v2.0.0'` now looks like this: ``` o e156455ea491 e156455ea491 gitster@pobox.com 2014-05-28 11:04:19.000 -07:00 refs/tags/v2.0.0 :\ Git 2.0 : ~ o c2f3bf071ee9 c2f3bf071ee9 junkio@cox.net 2005-12-21 00:01:00.000 -08:00 refs/tags/v1.0.0 ~ GIT 1.0.0 ``` Before this commit, it looked like this: ``` o e156455ea491 e156455ea491 gitster@pobox.com 2014-05-28 11:04:19.000 -07:00 refs/tags/v2.0.0 \| Git 2.0 \| o c2f3bf071ee9 c2f3bf071ee9 junkio@cox.net 2005-12-21 00:01:00.000 -08:00 refs/tags/v1.0.0 \| \|\ GIT 1.0.0 ``` The output of `jj log -r 'git_refs()'` in the git.git repo is still completely useless (it's >350k lines and >500MB of data). I think that's because we don't filter out edges to ancestors that we have transitive edges to. Mercurial also doesn't filter out such edges, but Git (with `--simplify-by-decoration`) seems to filter them out. I'll change it soon so we filter them out.	2021-05-01 14:56:52 -07:00
Martin von Zweigbergk	b953d7d801	git_store: revert lock timeout to 10s This backs out commit `67e11e0fc3`. We now use one thread per CPU in tests (`419002fab4`), so I hope the tests won't need the 1-minute timeout anymore.	2021-05-01 14:40:47 -07:00
Martin von Zweigbergk	df2caab274	tests: add a helper for building commit graphs when only topology is important	2021-04-30 22:46:20 -07:00
Martin von Zweigbergk	419002fab4	tests: use one thread per core in concurrency tests The tests have been failing in GitHub's CI quite frequently. It's about time I try to do something about it. Let's see if this helps.	2021-04-29 00:01:04 -07:00
Martin von Zweigbergk	46edbbef09	revsets: add revset function for getting all git refs This adds a `git_refs()` revset that includes all commits pointed to by a git ref. It's not very useful yet because the graph log doesn't use the right type of edges for non-contiguous commits.	2021-04-28 23:34:17 -07:00
Martin von Zweigbergk	f5151bdbbe	revsets: add a RevsetIterator type, to simplify API and enable adapters	2021-04-28 23:34:17 -07:00
Martin von Zweigbergk	0145be3693	index: add a newtype wrapper for IndexPosition It seems better to both hide the specific type and to get some more type safety.	2021-04-28 23:34:14 -07:00
Martin von Zweigbergk	5b18e89a4d	diff: fix LCS when a line/word/byte has been moved later	2021-04-28 23:33:18 -07:00
Martin von Zweigbergk	13134bd5a4	cleanup: address warnings reported by new clippy version	2021-04-28 09:12:48 -07:00
Martin von Zweigbergk	c6f6498cc9	conflicts: add newline after conflict marker lines Merging is currently done with line-level granularity, so it makes sense to have newlines after the markers. That makes them easier to edit out when resolving conflicts.	2021-04-24 13:53:24 -07:00
Martin von Zweigbergk	9dc18524fc	revsets: add "ancestor difference" range operator (like git's `..`)	2021-04-23 19:10:28 -07:00
Martin von Zweigbergk	49173de423	revsets: add DAG range operator (like hg's infix `::`) This lets you use the same operator as we currently have for ancestors and descendants (`,,`) to also specify a DAG range. That's what Mercurial uses the `::` operator for and what Git has `git log --ancestry-path` for.	2021-04-23 19:10:26 -07:00
Martin von Zweigbergk	d8c209c82a	revsets: give parents/children operators higher precedence than range operators	2021-04-23 18:45:42 -07:00
Martin von Zweigbergk	9de5f94af6	revsets: use same error variant for imcomplete parse as for syntax error	2021-04-23 16:59:04 -07:00
Martin von Zweigbergk	0b0374d401	revsets: make parsed Children and Descendants have roots and heads It seems clearer to let the parsed `RevsetExpression`s have only root and head expression instead of adding the ancestors when building the expression tree.	2021-04-23 16:15:17 -07:00
Martin von Zweigbergk	f209354503	revsets: simplify and clarify description revset slightly	2021-04-23 13:30:31 -07:00
Martin von Zweigbergk	145731ec74	revsets: change operators around a bit to prepare for infix DAG range operator I really liked the idea of having the operators for parents and ancestors (etc.) look similar, but that turned out to be problematic when we want to add an infix operator for a DAG range (hg's `::` revset operator and git's `--ancestry-path` flag). Let's say we chose `::` as the operator. Part of the problem is how to parse `foo::bar` without eagerly parsing the `foo:`. It would also be nicer to use exactly the same operator as prefix, postfix, and infix. Since the "parents" operator can be repeated, we can't have it be just `:` and the "ancestors" operator be `::`. We could make the "ancestors" operator be something like `:` (or anything symmetric with the `:` symbol on the inside). However, at that point, the operator is getting ugly and hard to type. Another option would be to use `:` for ancestors and `::` for parents, but that is counterintuitive and get annoying if you want to repeat it. So it seems that the best option is to simply pick different symbols for parents/children and ancestors/descendants/range. This patch changes the ancestors/descendants operators to both be `,,`. I'm not at all attached to that particular symbol. I suspect we'll change it later.	2021-04-23 11:11:07 -07:00
Martin von Zweigbergk	c894a7435f	revsets: make function arguments always be revset expressions Now that expressions may contain literal strings, we can simply have functions accept only expressions arguments. That simplifies both the grammar and the code. A small drawback is that `description((foo), bar)` is now allowed and does a search for the string "foo" (not "(foo)"). That seems unlikely to trip up users.	2021-04-23 10:51:41 -07:00
Martin von Zweigbergk	5819687237	revsets: accept quoted symbol names Git refs with names containing e.g "-" are currently not accepted symbol names, and I don't plan to change the grammar to accept them. Instead, let's have the user quote symbol names containing unusual characters. That way we can keep these symbols reserved for revset operators. With this patch the user can do e.g. `jj diff -r '"v2.9.0-rc2"'`.	2021-04-23 10:37:42 -07:00
Martin von Zweigbergk	d78fd9e979	revsets: add functions and operators for children and descendants This adds `children(<set>)` and `<set>:` for the children of the given set, and `descendants(<set>)` and `<set>:*` for the descendants of the given set. The children and descendants are filtered to be among ancestors of non-obsolete commits. I haven't added a way of overriding that yet.	2021-04-21 23:34:20 -07:00
Martin von Zweigbergk	618abf4379	revsets: use consistent "_op" suffix for operator rules in the grammar This is especially important now that we leak the rule names into the `SyntaxError` message. For example, the error message when doing `jj diff -r :` will now mention "expected parents_op, ancestors_op, or primary". It seems much clearer with the "_op" suffixes there. Longer term, we should think more about how we can best surface syntax errors from the library crate.	2021-04-21 18:53:58 -07:00
Martin von Zweigbergk	a4ef42962c	revsets: don't crash when given ungrammatical revset I also snuck in some updates to the test cases.	2021-04-21 18:53:58 -07:00
Martin von Zweigbergk	744f209e76	revsets: move parse-tests to to revset module (from separate test module) The tests don't need any complex set up (no repo necessary), so they can be in the `revset` module itself. I'm sure we'll need to split up that module later (at least separate out the parsing), but that's a separate problem.	2021-04-21 18:53:58 -07:00
Martin von Zweigbergk	6bc1361b84	index: make revision walk be by position instead of generation number I don't know why I made it walk by generation number to start with. Walking by position is better in at least two ways: 1) revsets now depend on the walks to be by descending index position (though they could equally well depend on the walks to be by generation number -- it just needs to be consistent), and 2) the log output gets less interleaved. This commit makes the number of bytes in the graphlog output in the git.git repo drop by ~40% due to the reduced amount of interleaving. Also, it reduces the time of `jj bench walkrevs v1.0.0 v2.0.0` in the git.git repo by 32% (9.4ms -> 6.4ms) and `jj bench walkrevs v2.0.0 v1.0.0` by 33% (7.7ms -> 5.1ms).	2021-04-21 18:53:56 -07:00
Martin von Zweigbergk	98f4e24892	cli: make benchmark ids include parameters It makes no sense to compare a run of `jj walkrevs v1.0.0 v2.0.0` with a run of `jj walkrevs v2.0.0 v1.0.0`, for example.	2021-04-21 16:56:45 -07:00
Martin von Zweigbergk	64fcf90c68	view: make root commit public	2021-04-18 23:04:15 -07:00
Martin von Zweigbergk	b52cfc156c	revsets: add a public_heads() revset function	2021-04-18 22:52:31 -07:00
Martin von Zweigbergk	3a65c1d2ab	revsets: add intersection operator	2021-04-18 22:45:12 -07:00
Martin von Zweigbergk	332580918c	revsets: add union operator	2021-04-18 22:45:12 -07:00
Martin von Zweigbergk	2ac5d1f912	revsets: allow spaces in most places (but not after prefix operators)	2021-04-18 22:45:12 -07:00
Martin von Zweigbergk	c04f418e67	revsets: add difference operator	2021-04-18 22:45:12 -07:00
Martin von Zweigbergk	e733b074e1	revsets: restructure grammar to prepare for operator precedences	2021-04-18 22:45:12 -07:00
Martin von Zweigbergk	d9ae7cdd6d	revsets: allow parenthesized expressions We'll clearly want to allow parenthesized expressions once we have infix operators (if not before). Let's prepare by allowing parentheses already now.	2021-04-18 22:45:12 -07:00
Martin von Zweigbergk	d71c083a7f	cli: use revsets also when looking up by description	2021-04-18 22:45:12 -07:00
Martin von Zweigbergk	62f0778942	revsets: add a description() revset The revset is currently eagerly evaluated, which is clearly bad. We'll need to fix that later.	2021-04-18 22:45:12 -07:00
Martin von Zweigbergk	05e9149157	revsets: add a non_obsolete_heads() revset This change adds a `non_obsolete_heads(<set>)` revset, which walks up ancestors of the input set until it gets to a non-obsolete and non-pruned commit. That's what we do by default in `jj log` (i.e. without `--all`). Now we can make `jj log` use revsets and teach it a `-r` option!	2021-04-18 22:45:10 -07:00
Martin von Zweigbergk	30aa459d2a	revsets: add a all_heads() revset function This adds a `all_heads()` revset function, which contains all heads in the view, i.e. including non-public heads and obsolete heads.	2021-04-18 22:31:46 -07:00
Martin von Zweigbergk	88904e2b63	revsets: add support for function syntax This adds `parents(foo)` and `ancestors(foo)` as alternative ways of writing `:foo` and `*:foo`. I haven't added support for for whitespace yet; the parsing is very strict. The error messages will also need to be improved later.	2021-04-18 21:25:58 -07:00
Martin von Zweigbergk	2d6325b0f4	revsets: define grammar in pest	2021-04-18 21:25:58 -07:00
Martin von Zweigbergk	0d62a336af	revsets: initial support for Mercurial-style revsets This patch adds initial support for a DSL for specifying revisions inspired by Mercurial's "revset" language. The initial support includes prefix operators ":" (parents) and ":" (ancestors) with naive parsing of the revsets. Mercurial uses postfix operator "^" for parent 1 just like Git does. It uses prefix operator "::" for ancestors and the same operator as postfix operator for descendants. I did it differently because I like the idea of using the same operator as prefix/postfix depending on desired direction, so I wanted to apply that to parents/children as well (and for predecessors/successors). The "" in the "*:" operator is copied from regular expression syntax. Let's see how it works out. This is an experimental VCS, after all. I've updated the CLI to use the new revset support. The implementation feels a little messy, but you have to start somewhere...	2021-04-18 21:25:51 -07:00
Martin von Zweigbergk	7861968f64	index: make IndexRef::entry_by_id() etc return entry with repo's lifetime It's useful to be able to know that given a `repo: RepoRef<'a>`, the the lifetime of `repo.index().entry_by_id()` will also be `'a`.	2021-04-15 07:00:04 -07:00
Martin von Zweigbergk	4c3d73ff3b	evolution: walk orphans using index This actually seems to make it slightly slower, but it fixes an important bug (we used to evolve only one topological branch per `jj evolve` call). The slowdown seemed to be on the order of 5% when evolving 100 commits on git.git's "what's cooking" branch.	2021-04-14 08:25:14 -07:00
Martin von Zweigbergk	783e1f6512	repo: make MutableRepo have an Arc<ReadonlyRepo> instead of a reference I suspect that at least one reason that I didn't make `MutableRepo::base_repo` by an `Arc<ReadonlyRepo>` before was that I thought that that would mean that `start_transaction()` would need be moved off of `ReadonlyRepo` so it can be given an `&Arc<ReadonlyRepo>`, which would make it much less convenient to use. It turns out that a `self` argument can actually be of type `&Arc<ReadonlyRepo>`.	2021-04-11 13:42:31 -07:00
Martin von Zweigbergk	ce855bccfa	repo: make reload() and reload_at() return a new ReadonlyRepo After this patch `ReadonlyRepo` is even closer to readonly. That makes it easier to reason about. It will allow some further cleanups too.	2021-04-11 10:39:29 -07:00
Martin von Zweigbergk	e3ca27bf77	revsets: support git refs	2021-04-10 10:10:09 -07:00
Martin von Zweigbergk	40f75ec641	revsets: don't crash if given non-hex symbol	2021-04-10 10:08:47 -07:00
Martin von Zweigbergk	9e8a7e2ba6	revsets: move code for resolving symbol to commit to new module	2021-04-10 09:46:27 -07:00
Martin von Zweigbergk	102f7a0416	diff: also recurse into final region after after unchanged regions See test case for details. Before: test bench_diff_10k_lines_reversed ... bench: 36,249,659 ns/iter (+/- 174,455) test bench_diff_10k_modified_lines ... bench: 37,258,890 ns/iter (+/- 803,963) test bench_diff_10k_unchanged_lines ... bench: 4,252 ns/iter (+/- 69) test bench_diff_1k_lines_reversed ... bench: 982,834 ns/iter (+/- 6,467) test bench_diff_1k_modified_lines ... bench: 3,343,469 ns/iter (+/- 23,243) test bench_diff_1k_unchanged_lines ... bench: 231 ns/iter (+/- 2) test bench_diff_git_git_read_tree_c ... bench: 95,559 ns/iter (+/- 816) After: test bench_diff_10k_lines_reversed ... bench: 36,186,715 ns/iter (+/- 196,903) test bench_diff_10k_modified_lines ... bench: 37,511,000 ns/iter (+/- 1,370,476) test bench_diff_10k_unchanged_lines ... bench: 3,099 ns/iter (+/- 8) test bench_diff_1k_lines_reversed ... bench: 986,010 ns/iter (+/- 11,565) test bench_diff_1k_modified_lines ... bench: 3,370,938 ns/iter (+/- 17,041) test bench_diff_1k_unchanged_lines ... bench: 230 ns/iter (+/- 2) test bench_diff_git_git_read_tree_c ... bench: 102,189 ns/iter (+/- 1,052) So this patch makes diffing even slower (but still easily fast enough for all cases I've run into in real life). There's probably a lot that can be done to make things faster, but the first priority is that the diffs are correct and easy to read.	2021-04-08 23:54:54 -07:00
Martin von Zweigbergk	f4a41f3880	trees: make tree diff return an iterator instead of taking a callback This is yet another step towards making it easy to propagate `BrokenPipe` errors. The `jj diff` code (naturally) diffs two trees and prints the diffs. If the printing fails, we shouldn't just crash like we do today. The new code is probably slower since it does more copying (the callback got references to the `FileRepoPath` and `TreeValue`). I hope that won't make a noticeable difference. At least `jj diff -r 334afbc76fbd --summary` didn't seem to get measurably slower.	2021-04-07 23:18:00 -07:00
Martin von Zweigbergk	8b2ce18254	trees: make diff_entries() return an iterator instead of taking a callback The iterator version is easier to use and we get rid of the ugly type parameter for the error type. I also simplified the code by using `Peekable` iterators.	2021-04-07 15:48:11 -07:00
Martin von Zweigbergk	5c10c93e64	diff: fix tests broken by the previous commit Sorry, I forgot to run the automated tests again :(	2021-04-07 11:00:04 -07:00
Martin von Zweigbergk	0dd000d236	diff: do final refinement at byte-level for non-word bytes This results in significantly more readable diffs on commits like `659393bec2` in this repo. Before: test bench_diff_10k_lines_reversed ... bench: 38,122,998 ns/iter (+/- 557,688) test bench_diff_10k_modified_lines ... bench: 32,556,563 ns/iter (+/- 548,114) test bench_diff_10k_unchanged_lines ... bench: 4,231 ns/iter (+/- 15) test bench_diff_1k_lines_reversed ... bench: 958,296 ns/iter (+/- 46,963) test bench_diff_1k_modified_lines ... bench: 3,014,723 ns/iter (+/- 15,830) test bench_diff_1k_unchanged_lines ... bench: 249 ns/iter (+/- 2) test bench_diff_git_git_read_tree_c ... bench: 78,599 ns/iter (+/- 1,079) After: test bench_diff_10k_lines_reversed ... bench: 38,289,493 ns/iter (+/- 413,712) test bench_diff_10k_modified_lines ... bench: 37,352,516 ns/iter (+/- 1,293,950) test bench_diff_10k_unchanged_lines ... bench: 4,238 ns/iter (+/- 13) test bench_diff_1k_lines_reversed ... bench: 967,253 ns/iter (+/- 8,506) test bench_diff_1k_modified_lines ... bench: 3,358,028 ns/iter (+/- 37,154) test bench_diff_1k_unchanged_lines ... bench: 233 ns/iter (+/- 1) test bench_diff_git_git_read_tree_c ... bench: 95,787 ns/iter (+/- 740) So the biggest slowdown is when there are modified lines.	2021-04-07 10:27:17 -07:00
Martin von Zweigbergk	f634ff0e3f	files: make diff() return an iterator instead of using a callback Iterators are generally nicer to work with. My immediate goal is to be able to propagate errors when failing to write to stdout.	2021-04-07 10:07:18 -07:00
Martin von Zweigbergk	d7395cc34a	diff: add copyright header	2021-04-06 21:26:37 -07:00
Martin von Zweigbergk	7e4e43f358	diff: first diff lines, then refine to words, producing better diffs The new diff algorithm produces pretty bad diffs in some cases, such as `cc4b1e9230` in this repo (the parent of this commit). I think the problem there is that many words are repeated over and over. Diffing first at the line level and then refining the diff of the changed ranges at the word level gives much better results. That's what this patch does. After this patch, `jj diff -r cc4b1e923091` looks pretty similar to the diff in GitHub's UI. I hope to get around to doing the same for the merge code soon. Impact on benchmarks: Before: test bench_diff_10k_lines_reversed ... bench: 42,647,532 ns/iter (+/- 765,347) test bench_diff_10k_modified_lines ... bench: 21,407,980 ns/iter (+/- 126,366) test bench_diff_10k_unchanged_lines ... bench: 4,235 ns/iter (+/- 16) test bench_diff_1k_lines_reversed ... bench: 1,190,483 ns/iter (+/- 7,192) test bench_diff_1k_modified_lines ... bench: 1,919,766 ns/iter (+/- 9,665) test bench_diff_1k_unchanged_lines ... bench: 231 ns/iter (+/- 1) test bench_diff_git_git_read_tree_c ... bench: 174,702 ns/iter (+/- 1,199) After: test bench_diff_10k_lines_reversed ... bench: 38,289,509 ns/iter (+/- 129,004) test bench_diff_10k_modified_lines ... bench: 33,140,659 ns/iter (+/- 3,989,339) test bench_diff_10k_unchanged_lines ... bench: 3,099 ns/iter (+/- 14) test bench_diff_1k_lines_reversed ... bench: 973,551 ns/iter (+/- 94,895) test bench_diff_1k_modified_lines ... bench: 3,033,818 ns/iter (+/- 29,513) test bench_diff_1k_unchanged_lines ... bench: 230 ns/iter (+/- 1) test bench_diff_git_git_read_tree_c ... bench: 79,100 ns/iter (+/- 963) So most of them get slower, as expected. The last one, taken from a real diff in the git.git repo, get faster, however (which is also what I would have expected).	2021-04-04 21:50:31 -07:00
Martin von Zweigbergk	cc4b1e9230	test: fix merge tests to expect line-based merging I made a quite late change in a recent patch to make the merge code to merge based on lines instead of words. I forgot to update the tests (and to even run them). Sorry :(	2021-04-01 08:27:27 -07:00
Martin von Zweigbergk	c071d412af	diff: use new diff algorithm for content diff The previous patch switched over the content-merge code to use the new histogram diff code. This patch switches over the content-diff code to use the histogram diff code. As before, the immediate goal is to speed it up. `jj diff -r c28ded83fc` in the git.git repo is a good example of a diff that's extremely slow to calculate with our current LCS-based diff. With this patch, that drops from 35 s to 0.12 s. The diff was slightly better before. I think that's mostly because of our different definition of a "word" in the data. We can improve that later. The speedup we get now is easily worth the slightly worse diff.	2021-03-31 22:22:59 -07:00
Martin von Zweigbergk	3c35dbace6	merge: use new diff algorithm for finding sync regions With the histogram diff code from the previous patch, we can now start using that for finding the "sync regions" in 3-way merge. That helps a lot with the slow merging we had before this patch. `jj diff -r 9d540e9726` in the git.git repo drops from 22 s to 0.15 s with this patch. (That commit is a rather arbitrary merge commit from aroun 5 years ago.) With the new diff algorithm, the output of `jj diff -r 9d540e9726` in git.git looks better if we find unchanged sync regions based on lines than on words, so that's what I'm using in this patch. That's a change compared the the LCS-based diff we used before this patch. I suspect the reason that finding sync regions based on words works worse now is not because of the change from LCS to histogram but because of the change in how we define a word. My goal right now is mostly to make it faster; I'll get back to refining the diff result later.	2021-03-31 22:16:19 -07:00
Martin von Zweigbergk	1e657c5331	diff: add a histogram(-like?) diff algorithm The current diff algorithm does a full LCS on the words of the texts, which is really slow. Diffing the working copy when e.g. `src/commands.py` has changes far apart takes seconds. This patch adds an implementation inspired by JGit's Histogram diff. I say "inspired" because I just didn't quite understand it :P In particular, I didn't understand what it does when it finds non-unique elements. I decided to line up the leading common elements on both sides of the merge. I don't know if that usually gives good enough results in practice. I'm sure this can still be optimized a lot, but this seems good enough as a start. There is also many things to improve about the quality of the diffs.	2021-03-31 22:15:36 -07:00
Martin von Zweigbergk	998e23db3c	index: add IndexEntry::parents() and predecessors() returning Vec<IndexEntry>	2021-03-31 14:48:03 -07:00
Martin von Zweigbergk	53d1757994	dag_walk: remove unused TopoIter	2021-03-18 16:42:30 -07:00
Martin von Zweigbergk	db4e8bc458	cargo: upgrade to protobuf 2.22.1 to avoid workaround for rustfmt::skip	2021-03-18 13:06:42 -07:00
Martin von Zweigbergk	07c2b2316f	repo: remove obsolete part of a TODO (we use the index to filter out non-heads)	2021-03-17 08:28:21 -07:00
Martin von Zweigbergk	30cd94f842	dag_walk: rename unreachable() to heads() to match name we use in index module	2021-03-16 23:54:51 -07:00
Martin von Zweigbergk	5aec8b9d77	evolution: use index for filtering out ancestors of candidates in new_parent() This speeds up `jj evolve` of 100 linear commits of the "what's cooking" branch in the git.git repo further, from ~700 ms to ~400 ms.	2021-03-16 23:43:44 -07:00
Martin von Zweigbergk	73f20c8696	transaction: delete write_commit() and as_repo_ref() helpers With this patch, the simple delegating helpers are gone from `Transaction`.	2021-03-16 22:45:58 -07:00
Martin von Zweigbergk	f9873c49ec	transaction: remove add_head(), remove_head(), and set_view() helpers	2021-03-16 22:31:28 -07:00
Martin von Zweigbergk	06df609482	transaction: delete check_out() and set_checkout() helpers	2021-03-16 22:31:28 -07:00
Martin von Zweigbergk	808d0af66d	transaction: remove evolution() and store() helpers	2021-03-16 22:31:24 -07:00
Martin von Zweigbergk	16d97ef8c0	transaction: remove index() and view() helpers	2021-03-16 22:05:51 -07:00
Martin von Zweigbergk	5ed14185a0	git: take a MutableRepo instead of a Transaction	2021-03-16 22:05:51 -07:00
Martin von Zweigbergk	769f88bbae	tests: rename test_transaction to test_mut_repo The test doesn't test any logic in the `Transaction` type itself anymore.	2021-03-16 22:05:51 -07:00
Martin von Zweigbergk	2c2b5fb3b7	evolution: take a MutableRepo instead of a Transaction	2021-03-16 22:05:51 -07:00
Martin von Zweigbergk	c3b9d1cd13	rewrite: take a MutableRepo instead of a Transaction	2021-03-16 22:05:51 -07:00
Martin von Zweigbergk	ee8423a69e	MutableRepo: rename `repo` to `base_repo` to clarify its role	2021-03-16 22:05:50 -07:00
Martin von Zweigbergk	69de4698ac	tests: set $HOME in a few tests to avoid depending in developer's ~/.gitignore I just changed my `~/.gitignore` and some tests started failing because the working copy respects the user's `~/.gitignore`. We should probably not depend on `$HOME` in the library crate. For now, this patch just makes sure we set it to an arbitrary directory in the tests where it matters.	2021-03-16 22:05:36 -07:00
Martin von Zweigbergk	67e11e0fc3	git_store: wait 1 minute for lock on refs to help tests `test_commit_parallel` was failing on Mac in the GitHub CI. I suspect the reason was that it was timing out. The test runs in about 1 s on my Linux desktop and in about 3 s on my Mac laptop. It failed after 31 in the GitHub CI. This patch increases the timeout to 1 minute to try to make the test pass. It would be better to set the timeout to a higher value only in tests, but this will be good enough for now. By the way, it has turned out that git notes (at least libgit2's implementation of them) are too slow, so we should probably eventually create our own storage for the extra metadata instead.	2021-03-16 11:28:22 -07:00
Martin von Zweigbergk	81a0e0bd2a	protobuf: upgrade to version 2.22.0 I only noticed that there was a newer version when running `cargo install --path .`, which resulted in warnings about deprecated functions. There's no other reason I'm aware of to upgrade now.	2021-03-15 17:09:29 -07:00
Martin von Zweigbergk	1ebdd4ecf0	MutableRepo: use index when enforcing view invariants We can now finally use the commit index for filtering out ancestors from the sets of heads. I haven't timed the change from most of the recent work on performance, but I did a measurement after this commit. I modified a commit in the git.git repo's "what's cooking" branch (because that's linear). Then I ran `jj evolve` so the 100 commits after it would get evolved. That took ~700ms. `git rebase` of the same 100 commits took ~6s. I also compared `jj op undo` of that `jj evolve` operation. With this patch, that was sped up from ~6.8s to ~125ms.	2021-03-15 16:35:45 -07:00
Martin von Zweigbergk	3ecb4ec16b	MutableRepo: in fast-path for adding head, simply remove parent heads	2021-03-15 15:38:09 -07:00
Martin von Zweigbergk	2c92fca75a	MutableView: don't require whole Commit when CommitId is enough	2021-03-15 15:36:03 -07:00
Martin von Zweigbergk	b4b1de3ddc	view: let MutableRepo enforce view invariants `MutableRepo` has more information needed for taking fast-paths, and it will have to make the same decision for doing incremental updates of the evolution state anyway.	2021-03-15 15:17:36 -07:00
Martin von Zweigbergk	b9fe944e76	view: remove unnecessary removing of parents in add_head() We call `enforce_invariants()` right after removing the parent commits, and that will remove parents anyway.	2021-03-15 15:06:14 -07:00
Martin von Zweigbergk	12a47bd6ed	MutableRepo: don't calculate evolution state only to update it	2021-03-15 15:03:50 -07:00
Martin von Zweigbergk	f0619c07ac	MutableEvolution: make MutableRepo responsible for lazy calculation This patch continues the work from the previous pathc. From this patch, we no longer calculate the evolution state just because a transaction starts. We still unnecessarily calculate it when adding a commit within the transaction, however. I'll fix that next.	2021-03-15 15:03:14 -07:00
Martin von Zweigbergk	61acee52f4	ReadonlyEvolution: make ReadonlyRepo responsible for lazy calculation This patch changes it so that `ReadonlyEvolution` does not lazily calculate its state and the caller, i.e. `ReadonlyRepo`, is instead responsible for the laziness. That will allow the caller to make decisions based on whether the state has been calculated. Specifically, we don't want to calculate the evolution state in order to update it incrementally if it hasn't already been calculated. It's better to just leave it uncalculated in that case. As a result of moving the laziness out of `ReadonlyEvolution`, we also don't need to the reference to `ReadonlyRepo` anymore, which simplifies things a bunch. The next patch will continue by making the corresponding change to `MutableEvolution`, which will let us simplify even more.	2021-03-15 14:41:27 -07:00
Martin von Zweigbergk	43315bc9d2	git: fix bad formatting from commit `1e9d428406`	2021-03-14 22:28:12 -07:00
Martin von Zweigbergk	91117f36b6	cargo: work around warning in generated protobuf code with new nightly rustc	2021-03-14 22:25:43 -07:00
Martin von Zweigbergk	1e9d428406	git: skip tags pointing to GPG keys and similar when importing refs	2021-03-14 20:14:18 -07:00
Martin von Zweigbergk	429a1ad7ab	git: set authentication callback on fetch as well I guess I had not run `jj git fetch` from GitHub until I tried to fetch the result of PR #6 just now.	2021-03-14 17:18:51 -07:00
Jun Wu	d1d502c062	tests: disable tests failing on Windows This unblocks enabling GitHub CI. I took a quick look at some failures but the causes do not seem obvious to me.	2021-03-14 15:51:32 -07:00
Jun Wu	935da3e13f	lock: treat PermissionDenied on Windows as transient error On Windows it can be PermissionDenied when creating the new file exclusively. This change makes lock_concurrent test pass on Windows.	2021-03-14 15:51:32 -07:00
Jun Wu	eacab648b0	working_copy: clean up ".git" automatically TreeState::write_tree leaves a ".git" file in the working copy. This is undesirable but more problematic on Windows - The second time TreeState::write_tree would panic because Repository::init_opts will fail with a Permission Denied error. This seems to be a libgit2 defect. But for now let's just remove ".git" automatically. This makes `cargo test --test smoke_test` pass on Windows.	2021-03-14 15:49:42 -07:00
Jun Wu	4cd29a2130	working_copy: avoid std::os::unix on Windows std::os::unix::fs::PermissionsExt::mode() does not exist on Windows. Treat files on Windows as regular files.	2021-03-14 15:49:22 -07:00
Martin von Zweigbergk	5631e85502	view: don't enforce invariants in merge_views() We now only call the function from `MutableRepo::merge()`. There we pass the result to `MutableView::set_view()`, which already enforces the invariants.	2021-03-14 11:07:34 -07:00
Martin von Zweigbergk	8048d9641e	commands: rewrite `jj op undo` using new `MutableRepo::merge()`	2021-03-14 10:57:57 -07:00
Martin von Zweigbergk	a7f4f4cf5b	rustfmt: configure to merge imports by module Perhaps we should even set the config to "Item" to reduce merge conflicts.	2021-03-14 10:53:14 -07:00
Martin von Zweigbergk	4b8484e561	rustfmt: configure to group imports	2021-03-14 10:46:25 -07:00
Martin von Zweigbergk	ac9fb1832d	OpHeadsStore: move logic for merging repos to MutableRepo This adds `MutableRepo::merge()`, which applies the difference between two `ReadonRepo`s to itself. That results in much simpler code than the current code in `merge_op_heads()`. It also lets us write `undo` using the new function. Finally -- and this is the actual reason I did it now -- it prepares for using the index when enforcing view invariants.	2021-03-14 10:43:39 -07:00
Martin von Zweigbergk	e9ddfdd8bc	Repo: repurpose ReadonlyRepo::loader() to return loader for existing repo It's sometimes useful to create a `RepoLoader` given an existing `ReadonlyRepo`. We already do that in `ReadonlyRepo::reload()`. This patch repurposes `ReadonlyRepo::reload()` for that.	2021-03-14 10:34:18 -07:00
Martin von Zweigbergk	82c683bf63	Transaction: rename as_repo_mut() to mut_repo() I think the `as_` prefix of `as_repo_mut()` makes it sound like it returns a view of the `Transaction`, but the `MutableRepo` is actually a part of it. Also, the convention seems to be to put the `mut_` in the name first if the function returns a name with a matching name (like `MutableRepo` does).	2021-03-14 00:25:05 -08:00
Martin von Zweigbergk	7ea0c6a868	View: move op_id/base_op_id to Repo This is yet another step towards making the `View` types simpler. Perhaps we eventually won't need to wrap the types returned from the `OpStore` at all.	2021-03-14 00:25:02 -08:00
Martin von Zweigbergk	c1de8b0f3a	View: move creation of Operation to Transaction This continues the work to make the `View` types be only about the state of the current view and not about operations in general (which has been moved out `OpStore` and qOpHeadsStore`).	2021-03-14 00:16:21 -08:00
Martin von Zweigbergk	cf2baf58a7	OpHeadsStore: simplify by returning Operation from get_single_op_head()	2021-03-14 00:16:21 -08:00
Martin von Zweigbergk	f6488e2e9f	OpHeadsStore: check for fast-forward merge before calling merge_op_heads() This is another little refactoring to prepare for using the `Transaction` API in `merge_op_heads()`.	2021-03-14 00:16:13 -08:00
Martin von Zweigbergk	9452d17b75	OpHeadsStore: pass around RepoLoader instead of various stores This is to prepare for using the regular `Transaction` API for creating the merge operation in `OpHeadStore`.	2021-03-14 00:15:04 -08:00
Martin von Zweigbergk	eac0c9f579	OpHeadsStore: when merging ops, also remove ancestor op from disk early This is more code, but I think it's clearer because the code for removing the ancestors from the set of parents and from disk is now close. I hope this will also help prepare for some further changes.	2021-03-14 00:13:04 -08:00
Martin von Zweigbergk	d4c39d399f	OpHeadsStore: read operation objects before calling merge_op_heads() This is just a little refactoring to prepare for filtering out ancestors earlier.	2021-03-14 00:13:04 -08:00
Martin von Zweigbergk	27293829d6	Transaction: allow writing a transaction to the OpStore without publishing it It can be useful to write an operation to the `OpStore` without also making it visible when you load the repo. I had planned to add that functionality at least for hooks, so the hooks can be run commands with `jj --at-op=<operation>` and decide whether to publish the operation. However, the immediate goal is to let us rewrite `op_heads_store::merge_op_heads()` to use the usual `Transaction` API. That needs to be able to just write the operation without publishing it, since the publishing step takes a long, which `op_heads_store::merge_op_heads()` (its caller, actually) has already taken.	2021-03-14 00:12:57 -08:00
Martin von Zweigbergk	337b15c98d	cleanup: replace #[cfg(not(windows))] by $[cfg(unix)] I didn't realize that the `unix` configuration existed before.	2021-03-12 15:45:55 -08:00
Martin von Zweigbergk	f79874d612	view: let repo get current operation from OpHeadsStore and pass in This makes the View types a lot simpler.	2021-03-11 22:23:02 -08:00
Martin von Zweigbergk	82a3ff6ef8	repo: make OpHeadsStore accessible directly on ReadonlyRepo We can now get rid of `MutableView::update_op_heads()`.	2021-03-10 23:27:36 -08:00
Martin von Zweigbergk	212dd35d01	view: let repo create OpHeadsStore and pass in to view	2021-03-10 23:14:00 -08:00
Martin von Zweigbergk	ec07104126	view: move creation of initial operation to OpHeadsStore This is a step towards getting rid of `MutableView::update_op_heads()`.	2021-03-10 22:57:59 -08:00
Martin von Zweigbergk	2590e127f7	view: move get_single_op_head() onto OpHeadsStore	2021-03-10 21:59:32 -08:00
Martin von Zweigbergk	b4d4cd143a	view: move locking of .jj/view/op_heads/ to OpHeadsStore	2021-03-10 21:51:08 -08:00
Martin von Zweigbergk	4bd121dab5	view: split out separate type for keeping track of op heads	2021-03-10 21:34:11 -08:00
Martin von Zweigbergk	2955bc4a29	repo: let repo types directly have an OpStore I'd like to make `ReadonlyView` and `MutableView` focused on just the state of the view (i.e. the set of heads, git refs, etc.). The responsibility for managing the `.jj/view/op_heads/` directory should be moved out of it. This prepares for that.	2021-03-10 20:55:56 -08:00
Martin von Zweigbergk	48d7903925	repo: simplify and clarify name of base_op_head_id() functions	2021-03-10 15:39:15 -08:00
Martin von Zweigbergk	9ee521d9d3	transaction: fix (mostly harmless) race where index can get re-calculated	2021-03-10 15:22:03 -08:00
Martin von Zweigbergk	47a7cf7101	view: extract function for updating operation heads This will be used to address the race in `Transaction::commit()`.	2021-03-10 15:17:54 -08:00
Martin von Zweigbergk	fc73ef8d6e	view: delete an incorrect comment about a race Unlike in `Transaction::commit()`, in the `view` module, we actually don't update the `.jj/view/op_heads/` directory until after we've recorded the index associated with the operation, so there's no race there.	2021-03-10 14:30:27 -08:00
Martin von Zweigbergk	a715fd0ae7	view: drop stale comment about resolving concurrent operations The comment was from the time when we resolved divergent operations at write time.	2021-03-08 23:18:26 -08:00
Martin von Zweigbergk	e6aa2402a6	view: drop redundant filtering of ancestors of public heads I added `enforce_invariants()` in `1f593a4193` and then forgot to use it in `4db3d8d3a6`.	2021-03-08 23:18:26 -08:00
Martin von Zweigbergk	f755c3f740	cleanup: access integer types' MAX constants directly on the type Using `std::u32::MAX` is deprecated.	2021-03-08 23:18:17 -08:00
Martin von Zweigbergk	02e6420606	repo: inline MutableRepo's {view,index,evolution}_mut() methods The methods are now only called from within the type. Inlining means that the borrow checker will let us borrow these separate fields concurrently. We'll take advantage of that soon.	2021-03-08 23:17:29 -08:00
Martin von Zweigbergk	9f7854f02c	repo: stop wrapping view and index in Option in MutableRepo I think the `Option<>` wrapping was from the time when `MutableView` had a reference back to the repo (and `MutableIndex` was probably wrapped out of habit).	2021-03-08 00:08:02 -08:00
Martin von Zweigbergk	ef16d102e2	transaction: move most functionality to MutableRepo Most methods on `Transaction` only need the `MutableRepo`, so it makes for that functionality to be on the latter. That will let us update the methods to also update the index, which would otherwise have been harder because it would require a mutable borrow of both the view and the index. This patch makes most current methods on `Transaction` just delegate to `MutableRepo`. We may want to remove some of these delegating methods later.	2021-03-07 23:10:32 -08:00
Martin von Zweigbergk	1e623bd019	index: update in memory and on disk while resolving operation conflicts Updating the index on disk means that reader won't have to calculate the state. Updating it in memory means that we can take advantage of it while resolving conflicts. We will do that soon.	2021-03-06 23:30:03 -08:00
Martin von Zweigbergk	779db67f8f	index_store: avoid passing whole repo into get_index_at_op() I want to be able to load the index at an operation before the repo has been loaded.	2021-03-06 23:06:35 -08:00
Martin von Zweigbergk	d1509ffdd4	index: extract a function for adding all commits from a segment I also restructured the loop in `maybe_squash_with_ancestors()` to hopefully make it a little clearer.	2021-03-06 20:30:12 -08:00
Martin von Zweigbergk	3fc35288c0	index: remove dir field from ReadonlyIndex and MutableIndex	2021-03-06 10:02:19 -08:00
Martin von Zweigbergk	502ba895f5	index: move ReadonlyFile::associate_with_operation() to IndexStore After this patch, the `index` module no longer knows about the ".jj/index/operations/" directory; that knowledge is now only in `IndexStore`.	2021-03-06 10:00:30 -08:00
Martin von Zweigbergk	c4fe7aab10	index: move ReadonlyRepo::load_at_operation() to IndexStore	2021-03-06 09:52:44 -08:00
Martin von Zweigbergk	12bfbc489c	index: move ReadonlyIndex::index() to IndexStore	2021-03-06 09:52:38 -08:00
Martin von Zweigbergk	2fdf9721c0	index: move load() to IndexStore	2021-03-06 09:52:27 -08:00
Martin von Zweigbergk	403e86c138	index: introduce IndexStore, which owns ReadonlyIndex files This patch introduces a new `IndexStore` struct. The idea is that it will know about the directory in which the index files are stored, the associations with operations. It may also cache `Arc<ReadonlyIndex>` instances so if multiple `ReadonlyIndex` instances are loaded, they can be returned from the cache. That may be useful when merging operations because the operations are likely to share a large parent index file. For now, however, all the new type has is `init()`, `load()`, and `reinit()`.	2021-03-06 09:52:16 -08:00
Martin von Zweigbergk	0a4ef1030f	repo: add support for loading at given operation without loading head op first The only way to load the repo at a current operation (as with `--at-op`) is currently to first load it at the head operation and then call `reload()` on the repo. This patch makes it so we can load the repo directly at the requested operation.	2021-03-06 09:52:10 -08:00
Martin von Zweigbergk	df53871daf	repo: extract a type for loading the repo in two stages We'll want to be able to load the repo at a given operation without first loading the head operation as we do today. This patch introduces a struct for keeping the state of a half-loaded repo. In that half-loaded state, the store and the op-store have been loaded, but the view has not yet been loaded. That makes it possible for callers to use the loaded op-store for looking up an operation to load the view at.	2021-03-06 09:52:10 -08:00
Martin von Zweigbergk	5a32118af1	repo: move creation of OpStore out of View We want to support loading the repo at a specific operation without first loading the head operation (like we currently do). One reason for that is of course efficiency. A possibly more important reason is that the head operation may be conflicted, depending on how we decide to deal with operation-level conflicts. In order to do that, it makes sense to move the creation of the `OpStore` outside of the `View`. That also suggests that the `.jj/view/op_store/` directory should move to `.jj/op_store/`, so this patch also does that. That's consistent with how `.jj/store/` is outside of `.jj/working_copy/`.	2021-03-06 09:52:00 -08:00
Martin von Zweigbergk	e2e9fe8f0d	index: add stats for number of change ids and pruned commits	2021-03-06 09:50:22 -08:00
Martin von Zweigbergk	bc64cf02c7	index: don't use an all-0 change id in tests It was weird to have the same change id for all commits. I think that was a leftover from a me just quickly getting tests to pass.	2021-03-06 09:30:52 -08:00
Martin von Zweigbergk	031a39ecba	cleanup: fix lots of issues found in the lib crate by clippy I had forgotten to pass `--workspace` to clippy all this time :P	2021-02-26 23:15:43 -08:00
Martin von Zweigbergk	d961f61623	evolution: calculate state using index All the information needed for calculating the evolution state is now in the index, so let's use it. This speeds up calculation of the evolution state from 1.53s to 150ms in the git.git repo. In the Linux repo, it was sped up from 28.9s to 3.07s. That's still unbearably slow (and still pretty slow in the git.git repo too). We may need to keep a persistent cache of the evolution state, but that will have to come later; this improvement is good enough for now.	2021-02-26 21:19:18 -08:00
Martin von Zweigbergk	190899fe76	index: make rev_walk() iterate over IndexEntry instead of CommitId We have the `IndexEntry` available in the iterator, so let's return it so the caller doesn't need to look it up themselves.	2021-02-26 16:49:27 -08:00
Martin von Zweigbergk	8be84c345b	index: also index change id	2021-02-26 10:33:34 -08:00
Martin von Zweigbergk	3a53a187ff	index: add flag indicating pruned commit	2021-02-26 10:33:34 -08:00
Martin von Zweigbergk	d80903ce48	index: also index predecessors Evolution needs to have fast access to the predecessors. This change adds that information to the commit index. Evolution also needs fast access to the change id and the bit saying whether a commit is pruned. We'll add those soon. Some tests changed because they previously added commits with predecessors that were not indexed, which is no longer allowed from this change. (We'll probably eventually want to allow that again, so that the user can prune predecessors they no longer care about from the repo.)	2021-02-26 10:33:34 -08:00
Martin von Zweigbergk	afc59a210a	index: make tests more focused, and add tests of octopus merge	2021-02-26 10:33:32 -08:00
Martin von Zweigbergk	2a531832d6	rewrite: make merge_commit_trees() use index for finding common ancestors The index is now always kept up to date and it has functionality for finding common ancestors, so let's use it! This should make merging commits a little faster if their common ancestor is far away (which is rare). It's probably much more important that the index-based algorithm is more correct. Also, it returns multiple common ancestors in the criss-cross case, which lets us do a recursive merge like git does. I'm leaving the recursive merge for later, though.	2021-02-23 20:49:18 -08:00
Martin von Zweigbergk	bb94516175	index: add support for finding common ancestors We currently need to read the commit objects for finding common ancestors. That can be very slow when the common ancestor is far back in history. This patch adds a function for finding common ancestors using the index instead. Unlike the current algorithm, which only returns one common ancestor, the new index-based one correctly handles criss-cross merges. Here are some timings for finding the common ancestors in the git.git repo: \| Without index \| With Index \| \| First run \| Subsequent \| First run \| Subsequent \| v2.30.0-rc0 v2.30.0-rc1 \| 5.68 ms \| 5.94 us \| 40.3 us \| 4.77 us \| v2.25.4 v2.26.1 \| 1.75 ms \| 1.42 us \| 13.8 ms \| 4.29 ms \| v1.0.0 v2.0.0 \| 492 ms \| 2.79 ms \| 23.4 ms \| 6.41 ms \| Finding ancestors of v2.25.4 and v2.26.1 got much slower because the new algorithm finds all common ancestors. Therefore, it also finds v2.24.2, v2.23.2, v2.22.3, v2.21.2, v2.20.3, v2.19.4, v2.18.3, and v2.17.4, which it then filters out because they're all ancestors of v2.25.3. Also note that the result was incorrect before, because the old algorithm would return as soon as it had found a common ancestor, even if it's not the latest common ancestor. For example, for the common ancestor between v1.0.0 and v2.0.0, it returned an ancestor of v1.0.0 because it happened to get there by following some side branch that led there more quickly. The only place we currently need to find the common ancestor is when merging trees, which we only do when the user runs `jj merge`, as well as when operating on existing merge commits (e.g. to diff or rebase them). That means that this change won't be very noticeable. However, it's something we clearly want to do sooner or later, so we might as well get it done.	2021-02-23 17:29:23 -08:00
Martin von Zweigbergk	422d333d4b	index: make heads() return result in index order instead of hash order It's nice to have a non-random order for tests (we can revisit later if it shows up in profiling). I'm changing the order to be the index order so the future caller of `heads_pos()` (not `heads()`) will also get consistent order.	2021-02-23 17:24:55 -08:00
Martin von Zweigbergk	1481935472	index: extract a function for removing ancestors of set based on positions We already have the `heads()` function, which works on `CommitId`s. This just extracts a function that works on positions. I'll use it soon.	2021-02-21 22:28:44 -08:00
Martin von Zweigbergk	5aadbcf6fc	evolve: pass Transaction to listener functions, so they see the updated state	2021-02-21 22:27:13 -08:00
Martin von Zweigbergk	62ce5782b5	index: when writing incremental index, squash into parent file if smaller We currently write a new incremental index file every time. That means that the stack of index files quickly gets deep, which makes it slow to read the index. This commit makes it so that we squash the new index segment into its parent if the parent has fewer commits. That means we'll limit the number of files to O(log n). Writes time will also be O(log n) on average.	2021-02-16 23:47:43 -08:00
Martin von Zweigbergk	a51543b752	index: make first level in stats be the root index I've confused myself a few times already thinking that level 0 is the root, so that's probably more intuitive. It also makes tests simpler because the initial part of the list is unchanged when a new transaction commits.	2021-02-16 23:45:54 -08:00
Martin von Zweigbergk	b122f33312	index: don't write empty incremental index file	2021-02-16 23:45:52 -08:00
Martin von Zweigbergk	a7b6bcfd79	transaction: write incremental index on commit With this change, we start writing the incremental index to disk, so the next reader won't have to re-read the commits and create the index. As of this change, we simply write a new index file for each transaction. That will clearly mean that the stack of files gets deep pretty quickly. For now, the user will have to do `jj debug reindex` when things get slow. I plan to change it so instead of writing an incremental index file every time, we first check if the new index file would have at least as many commits as the parent file, and if it will, we write a combined one instead. That should apply recursively, so we'd have O(log n) index files.	2021-02-15 11:03:41 -08:00
Martin von Zweigbergk	86915f0a6f	index: fix check for adding existing commit to index The check for adding an existing commit to the index only checked if the commit was already in the `MutableIndex`, not if it was already in the parent `ReadonlyIndex`.	2021-02-15 10:28:18 -08:00
Martin von Zweigbergk	37cf6a8395	transaction: don't walk to root when adding on top of non-head I don't know why I made the walk stop at heads instead of indexed commits before. Perhaps I did it because it's cheap to check in the set of head. However, it gets very expensive to walk all the way back to the root if the parents are not in the set of heads.	2021-02-15 10:28:18 -08:00
Martin von Zweigbergk	0f56e014b7	tests: some fixups to test_transaction as a result of reordering commits	2021-02-15 10:28:07 -08:00
Martin von Zweigbergk	3c832cbbbe	index: let index structs keep track of the index directory This matches how it's done for the other struct (View, WorkingCopy).	2021-02-14 01:03:49 -08:00
Martin von Zweigbergk	b77740e58a	index: move function for saving MutableIndex onto the struct	2021-02-14 01:03:49 -08:00
Martin von Zweigbergk	713d32d803	index: keep up to date within transaction With tons of groundwork done, wee can now finally keep the index up to date within a transaction! That means that we can start relying on the index to always be valid, so we can use it e.g. for finding common ancestors within a transaction. That should help speed up `jj evolve` immensely on large repos. We still don't write the updated index to disk when the transaction closes. That will come later.	2021-02-14 00:58:11 -08:00
Martin von Zweigbergk	e19a65cf14	transaction: make add_head() use incremental update of evolution in common case `Transaction::add_head()` currently invalidates the whole evolution state. We've had support for incrementally updating evolution since `4619942a57`. We should start taking advantage of that. Let's add a fast-path in `Transaction::add_head()` for the common case where we add a single commit on top of an existing head. That cheap an simple to check for. However, it won't cover the case of adding a child off of a non-head. It's still a good start.	2021-02-14 00:56:34 -08:00
Martin von Zweigbergk	f05a12d301	index: make CompositeIndex non-public and add new IndexRef enum instead We're getting close to finally having a `RepoRef::index()` method.	2021-02-13 13:56:26 -08:00
Martin von Zweigbergk	face4d637f	index: define methods from CompositeIndex directly on {Readonly,Mutable}Index This is one step towards making `CompositeIndex` non-public (and maybe deleting it). Next, we'll add an `IndexRef` enum similar to `RepoRef` etc.	2021-02-13 13:46:58 -08:00
Martin von Zweigbergk	8dda4b05e4	index: add "segment_" prefix to methods in IndexSegment I'm about to move the functions from `CompositeIndex` to an new `Index` trait implemeted by `ReadonlyIndex` and `MutableIndex`. Since those types already implement `IndexSegment`, the names would conflict and it would get annoying to have to disambiguate them. This commit therefore prepares for that by adding a `segment_` prefix to the functions in `IndexSegment`.	2021-02-13 13:45:31 -08:00
Martin von Zweigbergk	3066381d57	transaction: add accessors for view and evolution directly on transaction	2021-02-13 13:43:48 -08:00
Martin von Zweigbergk	72aebc9da3	view: replace View trait by enum with Readonly and Mutable variants	2021-02-13 08:31:41 -08:00
Martin von Zweigbergk	d1e5f46969	evolution: replace Evolution trait by enum with Readonly and Mutable variants	2021-02-13 08:31:41 -08:00
Martin von Zweigbergk	f1666375bd	repo: replace Repo trait by enum with readonly and mutable variants I want to keep the index updated within the transaction. I tried doing that by adding a `trait Index`, implemented by `ReadonlyIndex` and `MutableIndex`. However, `ReadonlyRepo::index` is of type `Mutex<Option<Arc<IndexFile>>>` (because it is lazily initialized), and we cannot get a `&dyn Index` that lives long enough to be returned from a `Repo::index()` from that. It seems the best solution is to instead create an `Index` enum (instead of a trait), with one readonly and one mutable variant. This commit starts the migration to that design by replacing the `Repo` trait by an enum. I never intended for there there to be more implementations of `Repo` than `ReadonlyRepo` and `MutableRepo` anyway.	2021-02-13 08:31:23 -08:00
Martin von Zweigbergk	a1983ebe96	git: add a ref to each commit we create I just learned that attaching a git note is not enough to keep a commit from being GC'd. I had read `git help gc` before but it was quite misleading (I just sent a patch to clarify it). Since the git note is not enough, we need to create some other reference. This patch makes it so we write refs in `refs/jj/keep/` for every commit we create. We will probably want to remove unnecessary refs (ancestors of commits pointed to by other refs) once we have a `jj gc` command.	2021-02-13 08:16:18 -08:00
Martin von Zweigbergk	dd98f0564e	git: remove git note pointing to conflicts We store conflicts as blobs with JSON data and with a git note pointing to them to prevent GC. These are stored in the git tree as regular files. The only thing that distinguishes them is that their filename ends with `.jjconflict`. Since they are referenced from the tree, there's no need for the git note to prevent GC (which doesn't work anyway, as I just learned), and we don't store any additional data in the note either, so let's just remove it.	2021-02-13 08:15:12 -08:00
Martin von Zweigbergk	fa30cf768f	index: rename UnsavedIndexData to MutableIndex	2021-02-07 23:35:37 -08:00
Martin von Zweigbergk	8170c06573	index: rename IndexFile to ReadonlyIndex	2021-02-07 23:35:22 -08:00
Martin von Zweigbergk	51373b75ff	index: use correct per-level file name in stats (previously always top-level)	2021-02-07 23:34:57 -08:00
Martin von Zweigbergk	302c66825f	working_copy: preserve executable bit on Windows Windows doesn't support recording the executable bit in the file system. Before this commit, the code for reading and writing the executable wouldn't even compile on Windows. This commit at least makes it so we preserve whatever bit has been recorded in the repo. At least I hope that's what it does -- I don't have access to a Windows machine right now.	2021-02-07 00:50:21 -08:00
Martin von Zweigbergk	d4aed83aa6	working_copy: correct comment about stat accuracy We don't take a write lock when writing a file; other processes can modify the file.	2021-02-07 00:48:50 -08:00
Martin von Zweigbergk	3d679de022	working_copy: print warning about ignored symlinks instead of failing build The project doesn't currently build on Windows. One reason is because we had a `unimplemented!()` when trying to write a symlink. Let's print a warning instead, so the project can start building on Windows. (The next patch will fix another build problem on Windows.)	2021-02-07 00:46:17 -08:00
Martin von Zweigbergk	e0112a4be0	transaction: report failure to close transaction only in debug builds When a transaction gets dropped without being committed or explicitly discarded, we currently raise an assertion error. I added that check because I kept forgetting to commit transactions. However, it's quite normal to want to drop transactions in error cases. The current assertion means that we panic and don't report the actual error to the user in such cases. We should probably audit the code paths where we commit transactions and decide for each if we simply want to to discard the transaction or not. In some cases, we may want to commit the transaction without integrating it in the operation log (i.e. without creating a file entry in .jj/views/op_heads/). However, we can do that later. For now, let's just make sure we don't panic when dropping the transaction in release builds.	2021-02-07 00:11:42 -08:00
Martin von Zweigbergk	4ecbd89378	repo: move MutableRepo from transaction module to repo module	2021-01-31 18:15:32 -08:00
Martin von Zweigbergk	2d03b514fc	transaction: move construction of MutableRepo out of Transaction::new() I'm about to move `MutableRepo` to the `repo` module and it will make more sense to have the construction of it there then.	2021-01-31 18:15:32 -08:00
Martin von Zweigbergk	bf53c6c506	transaction: add factory function to MutableRepo This helps finish the encapsulate of the `evolution` field.	2021-01-31 18:15:32 -08:00
Martin von Zweigbergk	5604303954	transaction: avoid direct access to members of MutableRepo I'm about to move `MutableRepo` to the `repo` module so it becomes more important to encapsulate access. Besides, the new functions introduced in this commit reduces some duplication. There's still one access of `MutableRepo::evolution` in `Transaction::new()`. I'll address that next by adding a factory function to `MutableRepo`.	2021-01-31 18:15:32 -08:00
Martin von Zweigbergk	a28fe7b388	transaction: slightly simplify write_commit() by using store()	2021-01-31 18:15:22 -08:00
Martin von Zweigbergk	9ffd35caf8	transaction: when checking out open commit with conflicts, create child commit I've been confused twice that rebasing an open commit so it results in conflicts doesn't show the conflicts in the log output. That's because we create a successor instead if a commit with conflicts is open. I guess I thought it would be expected that a child commit was not created. Since it seems surprising in practice, let's change it and we'll see if the new behavior is more or less surprising.	2021-01-22 11:41:52 -08:00
Martin von Zweigbergk	bb730d8a2b	merge: rewrite code for 3-way merge of files to handle not just trivial cases The most annoying remaining bug is that 3-way merge frequently panics with "unhandled merge case". This commit fixes that by rewriting the merge code. The new code is based on the algorithm used in Mercurial (which was in turn copied from Bazaar): 1. Find "sync" regions, which are regions that are the unchanged in the base and two sides. Note their start end end positions in each version. 2. Produce the output by taking the sync regions and inserting the result of merging the regions between the sync regions. These regions can either be changed on only one side, in which case we use that version, or it can be changed on both sides, in which case we indicate a conflict in the output. It's both more correct and much easier to follow.	2021-01-22 11:41:50 -08:00
Martin von Zweigbergk	7957feca49	diff: make tokenization return slices instead of making copies	2021-01-21 22:42:55 -08:00
Martin von Zweigbergk	30939ca686	view: return &HashSet instead of Iterator We want to be able to be able to do fast `.contains()` checks on the result, so `Iterator` was a bad type. We probably should hide the exact type (currently `HashSet` for both readonly and mutable views), but we can do that later. I actually thought I'd want to use `.contains()` for indiciting public-phase commits in the log output, but of course want to also indicate ancestors as public. This still seem like a step (mostly) in the right direction.	2021-01-16 13:00:05 -08:00
Martin von Zweigbergk	79eecb6119	git: mark imported remote-tracking branches as public	2021-01-16 12:14:42 -08:00
Martin von Zweigbergk	4db3d8d3a6	view: add tracking of "public" heads (copying Mercurial's phase concept) Mercurial's "phase" concept is important for evolution, and it's also useful for filtering out uninteresting commits from log output. Commits are typically marked "public" when they are pushed to a remote. The CLI prevents public commits from being rewritten. Public commits cannot be obsolete (even if they have a successor, they won't be considered obsolete like non-public commits would). This commits just makes space for tracking the public heads in the View.	2021-01-16 11:48:35 -08:00
Martin von Zweigbergk	265f90185e	tests: simplify transaction tests slightly by using testutils more	2021-01-16 11:31:57 -08:00

... 2 3 4 5 6 ...

411 commits