nix-super

mirror of https://github.com/privatevoid-net/nix-super.git synced 2024-11-24 23:06:16 +02:00

Author	SHA1	Message	Date
pennae	5d9fdab3de	use byte indexed locations for PosIdx we now keep not a table of all positions, but a table of all origins and their sizes. position indices are now direct pointers into the virtual concatenation of all parsed contents. this slightly reduces memory usage and time spent in the parser, at the cost of not being able to report positions if the total input size exceeds 4GiB. this limit is not unique to nix though, rustc and clang also limit their input to 4GiB (although at least clang refuses to process inputs that are larger, we will not). this new 4GiB limit probably will not cause any problems for quite a while, all of nixpkgs together is less than 100MiB in size and already needs over 700MiB of memory and multiple seconds just to parse. 4GiB worth of input will easily take multiple minutes and over 30GiB of memory without even evaluating anything. if problems do arise we can probably recover the old table-based system by adding some tracking to Pos::Origin (or increasing the size of PosIdx outright), but for time being this looks like more complexity than it's worth. since we now need to read the entire input again to determine the line/column of a position we'll make unsafeGetAttrPos slightly lazy: mostly the set it returns is only used to determine the file of origin of an attribute, not its exact location. the thunks do not add measurable runtime overhead. notably this change is necessary to allow changing the parser since apparently nothing supports nix's very idiosyncratic line ending choice of "anything goes", making it very hard to calculate line/column positions in the parser (while byte offsets are very easy).	2024-03-06 23:48:42 +01:00
pennae	855fd5a1bb	diagnose "unexpected EOF" at EOF this needs a string comparison because there seems to be no other way to get that information out of bison. usually the location info is going to be correct (pointing at a bad token), but since EOF isn't a token as such it'll be wrong in that this case. this hasn't shown up much so far because a single line ending is a token, so any file formatted in the usual manner (ie, ending in a line ending) would have its EOF position reported correctly.	2024-03-06 23:11:12 +01:00
pennae	1edd6fada5	report inherit attr errors at the duplicate name previously we reported the error at the beginning of the binding block (for plain inherits) or the beginning of the attr list (for inherit-from), effectively hiding where exactly the error happened. this also carries over to runtime positions of attributes in sets as reported by unsafeGetAttrPos. we're not worried about this changing observable eval behavior because it is marked unsafe, and the new behavior is much more useful.	2024-03-06 23:11:12 +01:00
pennae	cefd0302b5	evaluate inherit (from) exprs only once per directive desugaring inherit-from to syntactic duplication of the source expr also duplicates side effects of the source expr (such as trace calls) and expensive computations (such as derivationStrict).	2024-02-26 19:07:08 +01:00
pennae	c66ee57edc	preserve information about whether/how an attribute was inherited	2024-02-12 13:32:33 +01:00
Rebecca Turner	c0e7f50c1a	Rename `hintfmt` to `HintFmt`	2024-02-08 11:58:25 -08:00
Rebecca Turner	c6a89c1a16	libexpr: Support structured error classes While preparing PRs like #9753, I've had to change error messages in dozens of code paths. It would be nice if instead of EvalError("expected 'boolean' but found '%1%'", showType(v)) we could write TypeError(v, "boolean") or similar. Then, changing the error message could be a mechanical refactor with the compiler pointing out places the constructor needs to be changed, rather than the error-prone process of grepping through the codebase. Structured errors would also help prevent the "same" error from having multiple slightly different messages, and could be a first step towards error codes / an error index. This PR reworks the exception infrastructure in `libexpr` to support exception types with different constructor signatures than `BaseError`. Actually refactoring the exceptions to use structured data will come in a future PR (this one is big enough already, as it has to touch every exception in `libexpr`). The core design is in `eval-error.hh`. Generally, errors like this: state.error("'%s' is not a string", getAttrPathStr()) .debugThrow<TypeError>() are transformed like this: state.error<TypeError>("'%s' is not a string", getAttrPathStr()) .debugThrow() The type annotation has moved from `ErrorBuilder::debugThrow` to `EvalState::error`.	2024-02-01 16:39:38 -08:00
pennae	09a1128d9e	don't repeatedly look up ast internal symbols these symbols are used a lot, so it makes sense to cache them. this mostly increases clarity of the code (however clear one may wish to call the parser desugaring here), but it also provides a small performance benefit.	2024-01-15 16:52:18 +01:00
pennae	b596cc9e79	decouple parser and EvalState there's no reason the parser itself should be doing semantic analysis like bindVars. split this bit apart (retaining the previous name in EvalState) and have the parser really do only parsing, decoupled from EvalState.	2024-01-15 16:52:18 +01:00
pennae	e1aa585964	slim down parser.y most EvalState and Expr members defined here could be elsewhere, where they'd be easier to maintain (not being embedded in a file with arcane syntax) and somewhat more faithfully placed according to the path of the file they're defined in.	2024-01-15 16:52:18 +01:00
pennae	835a6c7bcf	rename ParserState::{makeCurPos -> at} most instances of this being used do not refer to the "current" position, sometimes not even to one reasonably close by. it could also be called `makePos` instead, but `at` seems clear in context.	2024-01-15 16:52:18 +01:00
pennae	0076056164	move ParseData to own header, rename to ParserState ParserState better describes what this struct really is. the parser really does modify its state (most notably position and symbol tables), so calling it that rather than obliquely "data" (which implies being input only) makes sense.	2024-01-15 16:52:18 +01:00
pennae	1b09b80afa	make parser utility functions members of ParseData all of them need access to parser state in some way. make them members to allow this without fussing so much.	2024-01-15 16:52:18 +01:00
pennae	e8d9de967f	simplify parse error reporting since nix doesn't use the bison `error` terminal anywhere any invocation of yyerror will immediately cause a failure. since we're already leaking tons of memory whatever little bit bison allocates internally doesn't much matter any more, and we'll be replacing the parser soon anyway. coincidentally this now also matches the error behavior of URIs when they are disabled or ~/ paths in pure eval mode, duplicate attr detection etc.	2024-01-15 16:52:18 +01:00
pennae	f07388bf98	remove ParserFormals this is a proper subset of Formals anyway, so let's just use those and avoid the extra allocations and moves.	2024-01-15 16:52:18 +01:00
John Ericson	e739a5002d	Avoid Windows macros in the parser and lexer `FLOAT`, `INT`, and `IN` are identifers taken by macros. The name `IN_KW` is chosen to match `OR_KW`, which is presumably named that way for the same reason of dodging macros.	2024-01-12 19:51:36 -05:00
pennae	b78e77b34c	use custom location type in the parser ~1% parser speedup from not using TLS indirections, less on system eval. this could have also gone in flex yyextra data, but that's significantly slower for some reason (albeit still faster than thread locals). before: Time (mean ± σ): 4.231 s ± 0.004 s [User: 3.725 s, System: 0.504 s] Range (min … max): 4.226 s … 4.240 s 10 runs after: Time (mean ± σ): 4.224 s ± 0.005 s [User: 3.711 s, System: 0.512 s] Range (min … max): 4.218 s … 4.234 s 10 runs	2023-12-19 19:32:16 +01:00
Eelco Dolstra	83c067c0fa	PosixSourceAccessor: Don't follow any symlinks All path components must not be symlinks now (so the user needs to call `resolveSymlinks()` when needed).	2023-12-05 23:02:59 +01:00
Eelco Dolstra	ea95327e72	Move restricted/pure-eval access control out of the evaluator and into the accessor	2023-11-30 16:16:17 +01:00
Eelco Dolstra	31ebc6028b	Fix symlink handling This restores the symlink handling behaviour prior to `94812cca98`. Fixes #9298.	2023-11-16 16:45:14 +01:00
John Ericson	ac89bb064a	Split up `util.{hh,cc}` All OS and IO operations should be moved out, leaving only some misc portable pure functions. This is useful to avoid copious CPP when doing things like Windows and Emscripten ports. Newly exposed functions to break cycles: - `restoreSignals` - `updateWindowSize`	2023-11-05 12:20:02 -05:00
Eelco Dolstra	955bbe53c5	Merge pull request #9177 from edolstra/input-accessors Backport FSInputAccessor and MemoryInputAccessor from lazy-trees	2023-10-23 11:42:04 +02:00
Eelco Dolstra	935c9981de	Remove fetchers::Tree and move tarball-related stuff into its own header	2023-10-20 19:56:52 +02:00
Eelco Dolstra	df73c6eb8c	Introduce MemoryInputAccessor and use it for corepkgs MemoryInputAccessor is an in-memory virtual filesystem that returns files like <nix/fetchurl.nix>. This removes the need for special hacks to handle those files.	2023-10-18 17:38:11 +02:00
Eelco Dolstra	ea38605d11	Introduce FSInputAccessor and use it Backported from the lazy-trees branch. Note that this doesn't yet use the access control features of FSInputAccessor.	2023-10-18 17:37:32 +02:00
Robert Hensing	16a6ea7249	Merge pull request #9049 from inclyc/users/inclyc/move-path libexpr: construct ExprPath by move ctor, not copy cotr	2023-09-27 22:30:44 +01:00
Yingchi Long	5b902ce9d6	libexpr: construct ExprPath by move ctor, not copy cotr	2023-09-26 23:30:32 +08:00
Robert Hensing	3720e811fa	libexpr: Add nrExprs to NIX_SHOW_STATS	2023-09-12 13:21:55 +02:00
John Ericson	fe71faa920	Delete `EvalState::addToSearchPath` This function is now trivial enough that it doesn't need to exist. `EvalState` can still be initialized with a custom search path, but we don't have a need to mutate the search path after it has been constructed, and I don't see why we would need to in the future. Fixes #8229	2023-08-18 14:04:33 -04:00
John Ericson	1570e80219	Move evaluator settings (type and global) to separate file/header	2023-07-31 10:14:15 -04:00
Naïm Favier	570a1a3ad7	parser: merge nested dynamic attributes Fixes https://github.com/NixOS/nix/issues/7115	2023-07-21 17:14:03 +02:00
John Ericson	4a880c3cc0	Merge pull request #8579 from obsidiansystems/findPath-cleanup-2 Further search path cleanups	2023-07-10 09:59:01 -04:00
Yingchi Long	3d74e7b811	libexpr: remove std::move() for `basePath` in parser, it has no effect	2023-07-10 12:02:29 +08:00
John Ericson	be518e73ae	Clean up `SearchPath` - Better types - Own header / C++ file pair - Test factored out methods - Pass parsed thing around more than strings Co-authored-by: Robert Hensing <roberth@users.noreply.github.com>	2023-07-09 23:22:22 -04:00
John Ericson	87dcd09047	Clean up `resolveSearchPathElem` We should use `std::optional<std::string>` not `std::pair<bool, std::string>` for an optional string.	2023-07-09 23:13:30 -04:00
Robert Hensing	1632f08ea2	Merge pull request #8600 from inclyc/libexpr/fix-leaking-in-stripIndentation libexpr: fix leaking `es2` in stripIndentation (parser.y)	2023-06-29 11:31:53 +02:00
Yingchi Long	3468cbaf47	libexpr: fix leaking `es2` in stripIndentation (parser.y)	2023-06-28 22:38:44 +08:00
John Ericson	484290a9e0	Use a struct not `std::pair` for `SearchPathElem` I got very confused trying to keep all the `first` and `second` straight reading the code, especially as there is also another `(boolean, string)` pair type also being used. Named fields is much better. There are other cleanups that we can do (for example, the existing TODO), but we can do them later. Doing them now would just make this harder to review.	2023-06-23 12:01:10 -04:00
Yingchi Long	9d8c4ac446	libexpr: remove unused token `ATTRPATH` in token declaration	2023-06-23 13:35:41 +08:00
Eelco Dolstra	1ad3328c5e	Allow tarball URLs to redirect to a lockable immutable URL Previously, for tarball flakes, we recorded the original URL of the tarball flake, rather than the URL to which it ultimately redirects. Thus, a flake URL like http://example.org/patchelf-latest.tar that redirects to http://example.org/patchelf-<revision>.tar was not really usable. We couldn't record the redirected URL, because sites like GitHub redirect to CDN URLs that we can't rely on to be stable. So now we use the redirected URL only if the server returns the `x-nix-is-immutable` or `x-amz-meta-nix-is-immutable` headers in its response.	2023-06-13 14:17:45 +02:00
Eelco Dolstra	a9759407e5	Origin: Use SourcePath	2023-04-06 15:25:06 +02:00
Eelco Dolstra	94812cca98	Backport SourcePath from the lazy-trees branch This introduces the SourcePath type from lazy-trees as an abstraction for accessing files from inputs that may not be materialized in the real filesystem (e.g. Git repositories). Currently, however, it's just a wrapper around CanonPath, so it shouldn't change any behaviour. (On lazy-trees, SourcePath is a <InputAccessor, CanonPath> tuple.)	2023-04-06 13:15:50 +02:00
John Ericson	296831f641	Move enabled experimental feature to libutil struct This is needed in subsequent commits to allow the settings and CLI args infrastructure itself to read this setting.	2023-03-20 11:05:22 -04:00
Eelco Dolstra	29abc8e764	Remove FormatOrString and remaining uses of format()	2023-03-02 15:57:54 +01:00
Et7f3	cec23f5dda	ExprOpHasAttr,ExprSelect,stripIndentation,binds,formals: delete losts objects We are looking for *$ because it indicate that it was constructed with a new but not release. De-referencing shallow copy so deleting as whole might create dangling pointer that's why we move it so we delete a empty containers + the nice perf boost.	2023-02-16 19:53:55 +01:00
Et7f3	fa89d317b7	ExprString: Avoid copy of string	2023-02-12 05:49:45 +01:00
Et7f3	3d16f2a281	parser: use implicit rule	2023-02-12 05:49:45 +01:00
Guillaume Maudoux	e4726a0c79	Revert "Revert "Merge pull request #6204 from layus/coerce-string"" This reverts commit `9b33ef3879`.	2023-01-19 13:23:04 +01:00
Robert Hensing	9b33ef3879	Revert "Merge pull request #6204 from layus/coerce-string" This reverts commit `a75b7ba30f`, reversing changes made to `9af16c5f74`.	2023-01-18 01:34:07 +01:00
Eelco Dolstra	6b69652385	Merge remote-tracking branch 'origin/master' into coerce-string	2023-01-02 20:53:39 +01:00

1 2 3 4 5 ...

300 commits