employer/tame - tame - Mike Gerwitz's Forge

employer

tame

Author	SHA1	Message	Date
Mike Gerwitz	13bac8382f	tamer: obj::xmlo::{air,reader}::test: Format test cases Simple reformatting that's consistent with other more recent tests, before I go making changes. DEV-13162	2023-05-05 10:26:58 -04:00
Mike Gerwitz	647e0ccbbd	tamer: Re-introduce literal parsing for xmlns in NIR for PackageStmt This _only_ re-introduces for PackageStmt since that's all I have tests for at present. More will be re-added later. They were previously removed when the attribute parsing was upended in `ele_parse!`. This does lose the attribute name, compared to before; that'll ideally be re-added, and I'll explore options for doing so later, since I also want them in other contexts. But it needs to be done generically (not XML-related). This had to be done before blowing up on TODOs, or system tests would fail. DEV-13708	2023-04-12 11:59:49 -04:00
Mike Gerwitz	f29e3cfce1	tamer: asg::air: Use StateStack This was extracted from xir::parse::ele in previous commits. The conventions help to ensure that pushing and returns are being performed correctly. The abstraction will continue to evolve. This ends up using `Ready` as the dead state. I need to determine if this is ideal, and if so, maybe just use `Default`, otherwise yield an error. DEV-13708	2023-03-30 15:44:14 -04:00
Mike Gerwitz	e6c6028b37	tamer: xir::parse::ele: Move StateStack into parse::state This will be utilized by `AirAggregate`. DEV-13708	2023-03-30 15:44:12 -04:00
Mike Gerwitz	11a4fdfb26	tamer: xir::parse::ele::StateStack: {Array=>}Vec The use of ArrayVec doesn't buy us anything anymore. There is no difference in performance through my own benchmarking (at least on our systems), and the game has changed since this was written: the size of the states is much smaller since we're no longer aggregating attributes. Further, the use of ArrayVec during development was also to keep memory allocation away from various parts of the code, which simplified analysis of the binary that was produced. Maybe it also reduced memory contention, but clearly that has no observable impact. The use of `Vec` removes the arbitrary bound, though I still kept one around just in case something goes wrong, so TAMER will terminate. Even though the token stream is bounded in size, lookahead does create recursion, and the system cannot (as written) prove that it doesn't. This is preparing for extracting `StateStack` into `parse` for use with `AirAggregate`. DEV-13708	2023-03-30 10:17:15 -04:00
Mike Gerwitz	e595698309	tamer: nir: Apply*Short variants This adds explicit variants for shorthand template application. This is less cryptic, and we'll be able to check for the close directly during desugaring. DEV-13708	2023-03-29 12:58:35 -04:00
Mike Gerwitz	9c0e20e58c	tamer: asg: Shorthand and long-form template arguments This applies to template application only; there's still some work to do for template parameters in definitions (well, for deriving them in `xmli` at least). And, as you can see, there's still a lot of TODO items here. I ended up backtracking on tree edges to Meta, and even on cross edges to Meta, because it complicated xmli derivation with no benefit right now; maybe a cross edge will be re-added in the future, but I need to move on and see where this takes me. But, it works. DEV-13708	2023-03-29 12:58:35 -04:00
Mike Gerwitz	bef68e1634	tamer: nir: Desugar shorthand template params and yield AIR I had intended for this to be a full vertical slice initially, but AIR's parser is going to need enough work that it'll muddy this patch a bit too much. This keeps the desugaring simple, which is what I was hoping for. The next step is to load it into the graph and emit regenerated longhand sources. I also don't like how the namespace prefix is just being ignored for shorthand param desugaring. This is also the case in the XSLT-based compiler, but this violates TAMER's principle that it should parse every bit of information; nothing should be ignored. If something does not contribute useful information, then it is not a useful construct and ought to be rejected. DEV-13708	2023-03-29 12:58:35 -04:00
Mike Gerwitz	aec3b97e3f	tamer: parse::parser::Parser: Prevent infinite iteration on finalize This was a rather frustrating thing to encounter. I was working on refactoring `AirAggregate`, and found that my tests were hanging despite no apparent cause in the parser itself. As it turns out, rather than failing with a `FinalizeError` as I expected (since I was mid-refactor), `collect()` was allocating space for an endless stream of errors. This was easily verified by adding a `take(x)` and observing the assertion failure (in this case, in `close_pkg_mid_expr`). This happens to be the first time in a long time that I actually had to debug---the combination of robust types as proofs and tests to fill in the gaps means that runtime issues are caught at build time in all but exceptional cases (like this one). It's also worth noting that, because of my policy of iterating only at the higher levels of the program, it was clear that this must somehow be Parser-related, since that's the only part of the system that has the potential for unbounded recursion due to its cyclic state machines. DEV-13708	2023-03-10 14:27:58 -05:00
Mike Gerwitz	82915f11af	tamer: asg::graph::object::xir: Initial rate element reconstruction This extends the POC a bit by beginning to reconstruct rate blocks (note that NIR isn't producing sub-expressions yet). Importantly, this also adds the first system tests, now that we have an end-to-end system. This not only gives me confidence that the system is producing the expected output, but serves as a compromise: writing unit or integration tests for this program derivation would be a great deal of work, and wouldn't even catch the bugs I'm worried most about; the lowering operation can be written in such a way as to give me high confidence in its correctness without those more granular tests, or in conjunction with unit or integration tests for a smaller portion. DEV-13708	2023-03-10 14:27:58 -05:00
Mike Gerwitz	6db70385d0	tamer: xir::flat: Introduce configurable acceptors Technically, an "acceptor" in the context of state machines is actually a state machine; the terminology here is more describing the configuration of the state machine (`XirToXirf`) as an acceptor. This change comes with significant documentation of the rationale and why this is important; see that for more information. This change is necessary so that we can enforce finalization on all parsers in the lowering pipeline, which is not currently being done. If we were to do that now, then `tameld` would fail because it halts parsing of the tokens stream at the end of the `xmlo` header. This is also quite the type soup, but I'm not going to refine this further right now, since my focus is elsewhere (XMLI lowering). DEV-13708	2023-03-10 14:27:57 -05:00
Mike Gerwitz	33d2b4f0b8	tamer: tamec: POC lowering pipeline with XirfAutoClose and XirfToXir This replaces the stub `derive_xmli` with the same result (well, minus a space before the '/' in the output) using what will become the lowering pipeline. Once again, this is quite verbose, and the lowering pipeline in general needs to be further abstracted away. Unlike the rest of the pipeline, an error during the derivation process will immediately terminate with an unrecoverable error, because we do not want to write partial files. This does not remove the garbage file, because the build system ought to do that itself (e.g. `make`)...but that is certainly open for debate. DEV-13708	2023-03-10 14:27:57 -05:00
Mike Gerwitz	29178f2360	tamer: xir::reader: Divorce from `parse` The reader previously yielded a `ParsedResult`, presumably to simplify lowering operations. But the reader is not a `ParseState`, and does not otherwise use the parsing API, so this was an inappropriate and confusing coupling. This resolves that, introducing a new `lowerable` which will translate an iterator into something that can be placed in a lowering pipeline. See the previous commit for more information. DEV-13708	2023-03-10 14:27:57 -05:00
Mike Gerwitz	963688f889	tamer: parse::lower::ParsedObject: Include Token type parameter The token type was previously hard-coded to `UnknownToken`, since the use case was the beginning of the lowering pipeline at the start of the program, where there was no token type because the first parser (`XirReader`, currently) is responsible for producing the first token type. But when we're lowering from the graph (so, the other side of the lowering pipeline), we _do_ have token types to deal with. This also emphasizes the inappropriate coupling of `<XirReader as Iterator>::Item` with `ParsedResult`; I'd like to follow the same approach that I'm about to introduce with `tamec`, so see a future commit. DEV-13708	2023-03-10 14:27:57 -05:00
Mike Gerwitz	79cc61f996	tamer: xir::flat::XirfToXir: New lowering operation This parser does exactly what it says it does. Its implementation is simple, but I added a test anyway just to prove that it works, and the test seems more complicated than the implementation itself, given the types involved. DEV-13708	2023-03-10 14:27:57 -05:00
Mike Gerwitz	bc8586e4b3	tamer: xir::autoclose: New lowering operation This lowering operation is intended to allow me to write a more concise and clear mapping from the graph to XIRF, without having to worry about balancing tags, which really complicated the implementation. This has details docs; see that for more information. I can't help but be reminded of Wisp (the whitespace-based Lisp-like syntax). Which is unfortunate, because I'm not fond of Wisp; I like my parenthesis. DEV-13708	2023-03-10 14:27:57 -05:00
Mike Gerwitz	065dca88fc	tamer: asg::graph::vist::tree_reconstruction: Include Depth This information is necessary to be able to reconstruct the tree, since the `ObjectIndex` alone does not give you enough information. Even if you inspected the graph, it _still_ wouldn't give you enough information, since you don't know the current path of the traversal for nodes that may have multiple incoming edges. (Any assumptions you could make today won't always be valid in the future.) DEV-13708	2023-03-10 14:27:57 -05:00
Mike Gerwitz	954b5a2795	Copyright year and name update Ryan Specialty Group (RSG) rebranded to Ryan Specialty after its IPO.	2023-01-20 23:37:30 -05:00
Mike Gerwitz	e6640c0019	tamer: Integrate clippy This invokes clippy as part of `make check` now, which I had previously avoided doing (I'll elaborate on that below). This commit represents the changes needed to resolve all the warnings presented by clippy. Many changes have been made where I find the lints to be useful and agreeable, but there are a number of lints, rationalized in `src/lib.rs`, where I found the lints to be disagreeable. I have provided rationale, primarily for those wondering why I desire to deviate from the default lints, though it does feel backward to rationalize why certain lints ought to be applied (the reverse should be true). With that said, this did catch some legitimage issues, and it was also helpful in getting some older code up-to-date with new language additions that perhaps I used in new code but hadn't gone back and updated old code for. My goal was to get clippy working without errors so that, in the future, when others get into TAMER and are still getting used to Rust, clippy is able to help guide them in the right direction. One of the reasons I went without clippy for so long (though I admittedly forgot I wasn't using it for a period of time) was because there were a number of suggestions that I found disagreeable, and I didn't take the time to go through them and determine what I wanted to follow. Furthermore, it was hard to make that judgment when I was new to the language and lacked the necessary experience to do so. One thing I would like to comment further on is the use of `format!` with `expect`, which is also what the diagnostic system convenience methods do (which clippy does not cover). Because of all the work I've done trying to understand Rust and looking at disassemblies and seeing what it optimizes, I falsely assumed that Rust would convert such things into conditionals in my otherwise-pure code...but apparently that's not the case, when `format!` is involved. I noticed that, after making the suggested fix with `get_ident`, Rust proceeded to then inline it into each call site and then apply further optimizations. It was also previously invoking the thread lock (for the interner) unconditionally and invoking the `Display` implementation. That is not at all what I intended for, despite knowing the eager semantics of function calls in Rust. Anyway, possibly more to come on that, I'm just tired of typing and need to move on. I'll be returning to investigate further diagnostic messages soon.	2023-01-20 23:37:29 -05:00
Mike Gerwitz	9103c93693	tamer: xir::writer: write{=>_all} Really, with a C background, I should have known that `write` may not write all bytes, and I'm pretty sure I was aware, so I'm not sure how that slipped my mind for every call. But it's not a great default, and I do feel like `write_all` should be the deafult behavior, despite the syscall and C library name. It shouldn't take clippy to warn about something so significant.	2023-01-01 23:43:00 -05:00
Mike Gerwitz	07dff3ba4e	tamer: xir::parse::ele: Remove attr sum state This removes quite a bit of work, and work that was difficult to reason about. While I'm disappointed that that hard work is lost (aside from digging it up in the commit history), I am happy that it was able to be removed, because the extra complexity and cognitive burden was significant. This removes more `memcpy`s than the sum state could have hoped to, since aggregation is no longer necessary. Given that, there is a slight performacne improvement. The re-introduction of required and duplicate checks later on should be more efficient than this was, and so this should be a net win overall in the end. DEV-13346	2022-12-01 11:09:26 -05:00
Mike Gerwitz	f872181f64	tamer: xir::parse: Remove old `attr_parse!` and unused error variants This cleans up the old implementation now that it's no longer used (as of the previous commit) by `ele_parse!`. It also removes the two error variants that no longer apply: required attributes and duplicate attributes. DEV-13346	2022-12-01 11:09:26 -05:00
Mike Gerwitz	ab0e4151a1	tamer: xir::parse::ele::ele_parse!: Integrate `attr_parse_stream!` This handles the bulk of the integration of the new `attr_parse_stream!` as a replacement for `attr_parse!`, which moves from aggregate attribute objects to a stream of attribute-derived tokens. Rationale for this change is in the preceding commit messages. The first striking change here is how it affects the test cases: nearly all `Incomplete`s are removed. Note that the parser has an existing optimization whereby `Incomplete` with lookahead causes immediate recursion within `Parser`, since those situations are used only for control flow and to keep recursion out of `ParseState`s. Next: this removes types from `nir::parse`'s grammar for attributes. The types will instead be derived from NIR tokens later in the lowering pipeline. This simplifies NIR considerably, since adding types into the mix at this point was taking an already really complex lowering phase and making it ever more difficult to reason about and get everything working together the way that I needed. Because of `attr_parse_stream!`, there are no more required attribute checks. Those will be handled later in the lowering pipeline, if they're actually needed in context, with possibly one exception: namespace declarations. Those are really part of the document and they ought to be handled _earlier_ in the pipeline; I'll do that at some point. It's not required for compilation; it's just required to maintain compliance with the XML spec. We also lose checks for duplicate attributes. This is also something that ought to be handled at the document level, and so earlier in the pipeline, since XML cares, not us---if we get a duplicate attribute that results in an extra NIR token, then the next parser will error out, since it has to check for those things anyway. A bunch of cleanup and simplification is still needed; I want to get the initial integration committed first. It's a shame I'm getting rid of so much work, but this is the right approach, and results in a much simpler system. DEV-13346	2022-12-01 11:09:26 -05:00
Mike Gerwitz	1983e73c81	tamer: xir::parse::attrstream: Value from SPair This really does need documentation. With that said, this changes things up a bit: the value is now derived from an `SPair` rather than an `Attr`, given that the name is redundant. We do not need the attribute name span, since the philosophy is that we're stripping the document and it should no longer be important beyond the current context. It does call into question errors, but my intent in the future is to be able to have the lowering pipline augment errors with its current state---since we're streaming, then an error that is encountered during lowering of an element will still have the element parser in the state representing the parsing of that element; so that information does not need to be propagated down the pipeline, but can be augmented as it bubbles back up. More on that at some point in the future; not right now. DEV-13346	2022-12-01 11:09:25 -05:00
Mike Gerwitz	9ad7742ad2	tamer: xir::parse::attrstream: Streaming attribute parser As I talked about in the previous commit, this is going to be the replacement for the aggreagte `attr_parse!`; the next commit will integrate it into `ele_parse!` so that I can begin to remove the old one. It is disappointing, since I did put a bit of work into this and I think the end result was pretty neat, even if was never fully utilized. But, this simplifies things significantly; no use in maintaining features that serve no purpose but to confound people. DEV-13346	2022-12-01 11:09:25 -05:00
Mike Gerwitz	76beb117f9	Revert "tamer: nir::desugar::interp: Include attribute name in derived param name" Also: Revert "tamer: nir::desugar::interp: Token {SPair=>Attr}" This reverts commit 7fd60d6cdafaedc19642a3f10dfddfa7c7ae8f53. This reverts commit 12a008c66414c3d628097e503a98c80687e3c088. This has been quite a tortured experience, trying to figure out how to best fit desugaring into the existing system. The truth is that it ultimately failed because I was not sticking with my intuition---I was trying to get things out quickly by compromising on the design, and in the end, it saved me nothing. But I wouldn't say that it was a waste of time---the path was a dead end, but it was full of experiences. More to come, but interpolation is back to operating on NIR directly, and I chose to treat it as a source-to-source mapping and not represent it using the type system---interpolation can be an optional feature when writing TAME frontends (the principal one being the XML-based one), and it's up to later checks to assert that identifiers match a given domain. I am disappointed by the additional context we lose here, but that can always be introduced in the future differently, e.g. by maintaining a dictionary of additional context for spans that can be later referenced for diagnostic purposes. But let's worry about that in the future; it doesn't make sense to further complicate IRs for such a thing. DEV-13346	2022-12-01 11:09:25 -05:00
Mike Gerwitz	d0a728c27f	tamer: nir::desugar::interp: Token {SPair=>Attr} This changes the input token from a more generic `SPair` to `Attr`, which reflects the new target integration point in the `attr_parse!` parser-generator. This is a compromise---I'd like for it to remain generic and have stitching deal with all integration concerns, but I have spent far too much time on this and need to keep moving. With that said, we do benefit from knowing where this must fit in---it's easier to reason about in a more concrete way, and we can take advantage of the extra information rather than being burdened by its presence and ignoring it. We need to be able to convert back into `XirfToken` (see a recent commit that discusses that) for `StitchExpansion`, which is why `Attr` is here. And since it is, we can use it to explain to the user not just the interpolation specification used to derive params, but also the attribute it is associated with. This is what TAME (in XSLT) does today, IIRC (I wrote it, I just forget exactly). It also means that I can name the parameters after the attribute. So, that'll be in a following commit; I was disappointed when my prior approach with `SPair` didn't give me enough information to be able to do that, since I think it's important that the system be as descriptive as possible in how it derives information. Of course, traces would reveal how the parser came about the derivation, but that requires recompilation in a special tracing mode. DEV-13156	2022-12-01 11:09:25 -05:00
Mike Gerwitz	8a430a52bc	tamer: xir::prase: Extract intermediate attribute aggregate state into Context This was a substantial change. Design and rationale are documented on `AttrFieldSum` and related as part of this change, so please review the diff for more information there. If you're a Ryan employee, DEV-13209 gives plenty of profiling information, including raw data and visualizations from kcachegrind. For everyone else: you're able to easy produce your own from this commit and the previous and comparing the `__memcpy_avk_unaligned_erms` calls. The reduction is significant in this commit (~90%), and the number of Parsers invoking it has been reduced. Rust has been able to optimize more aggressively, and compound some of those optimizations, with the smaller `NirParseState` width. It also worth noting that `malloc` calls do not change at all between these two changes, so when we refer to memory, we're referring to pre-allocated memory on the stack, as TAMER was designed to utilize. DEV-13209	2022-11-09 16:01:09 -05:00
Mike Gerwitz	d195eedacb	tamer: nir: Sugared and plain flavors This introduces the concept of sugared NIR and provides the boilerplate for a desugaring pass. The earlier commits dealing with cleaning up the lowering pipeline were to support this work, in particular to ensure that reporting and recovery properly applied to this lowering operation without adding a ton more boilerplate. DEV-13158	2022-10-26 14:19:19 -04:00
Mike Gerwitz	26aaf6efc1	tamer: parse::error::ParseError: Extract some variants into FinalizeError This helps to clarify the situations under which these errors can occur, and the generality also helps to show why the inner types are as they are (e.g. use of `String`). But more importantly, this allows for an error type in `finalize` that is detached from the `ParseState`, which will be able to be utilized in the lowering pipeline as a more general error distinguishable from other lowering errors. At the moment I'm maintaining BC, but a following commit will demonstrate the use case to introduce recoverable vs. non-recoverable errors. DEV-13158	2022-10-26 12:44:19 -04:00
Brandon Ellis	00f46b0032	[DEV-12990] Add gt, gte, lt, lte operators to if/unless This includes updating Tamer's parser to account for the new operator possibilities.	2022-09-22 11:38:06 -04:00
Mike Gerwitz	dcb42b6e4b	tamer: xir::parse: Improvements to generated docs for NIR attributes This hides the internal state machine and provides better language for what remains. DEV-7145	2022-09-16 13:37:46 -04:00
Mike Gerwitz	f9bdcc2775	tamer: xir::parse::ele: Remove `*Error_` types A type alias was added for BC before errors were hoisted out in a previous commit, but they are unnecessary because of the associated type on `ParseState`. This also corrects the long-existing issue of using generated identifiers in tests. DEV-7145	2022-09-15 16:10:47 -04:00
Mike Gerwitz	071c94790f	tamer: xir::ele::parse: Formatting: remove a level of indentation This moves `paste::paste!` up a line and reduces a level of indentation, since it's so squished. Aside from docblock reformatting, there are no other changes. DEV-7145	2022-09-15 16:09:49 -04:00
Mike Gerwitz	b3f4378517	tamer: xir::parse::ele: Hoist NT Display from `ele_parse!` macro This slims out the macro even further. It does result in an awkwardly-placed `PhantomData` because I don't want to add another variant that isn't actually used (since they represent states). DEV-7145	2022-09-14 16:34:59 -04:00
Mike Gerwitz	80f29e9420	tamer: xir::parse::ele: Hoist NtState out of `ele_parse!` macro This does the same as before with SumNtState, and takes advantage of the preparations made by the preceding commit. The macro is shrinking. DEV-7145	2022-09-14 15:35:58 -04:00
Mike Gerwitz	1817659811	tamer: xir::parse::ele: Abstract child NT states in parent parser This is in preparation for hoisting out the common states, as was done with the Sum NT in a previous commit. I also think that organizing states in this way is more clear. The previous embedding of the variants named after the NTs themselves was because the parser was storing the child state within it, before the introduction of the superstate trampoline. DEV-7145	2022-09-14 14:47:54 -04:00
Mike Gerwitz	d73a18d1a2	tamer: xir::parse::ele: Initial extraction of Sum NT state from macro After introducing the superstate and trampoline some time ago, the Sum NT states became fully generalized and can be hoisted out. DEV-7145	2022-09-14 12:23:52 -04:00
Mike Gerwitz	db3fd3f177	tamer: xir::parse::ele: Remove `unreachable!` in state transitions This will instead fail at compile time. DEV-7145	2022-09-14 10:00:10 -04:00
Mike Gerwitz	a5c7067c68	tamer: xir::parse::ele: Remove NT `todo!` for state transition Everything except for one state was already accounted for. We can now have confidence that the parser will never panic due to state transitions (beyond legitimate error conditions). There are some `unreachable!`s to contend with still. DEV-7145	2022-09-14 09:41:53 -04:00
Mike Gerwitz	212ca06efe	tamer: xir::parse: Extract and generalize NT errors This is the same as the previous commits, but for non-sum NTs. This also extracts errors into a separate module, which I had hoped to do in a separate commit, but it's not worth separating them. My _original_ reason for doing so was debugging (I'll get into that below), but I had wanted to trim down `ele.rs` anyway, since that mess is large and a lot to grok. My debugging was trying to figure out why Rust was failing to derive `PartialEq` on `NtError` because of `AttrParseError`. As it turns out, `AttrParseError::InvalidValue` was failing, thus the introduction of the `PartialEq` trait bound on `AttrParseState::ValueError`. Figuring this out required implementing `PartialEq` myself without `derive` (well, using LSP, which did all the work for me). I'm not sure why this was not failing previously, which is a bit of a concern, though perhaps in the context of the macro-expanded code, Rust was able to properly resolve the types. DEV-7145	2022-09-14 09:28:31 -04:00
Mike Gerwitz	5078bd8bda	tamer: xir::parse::ele: Extract sum NT error from `ele_parse!` The `ele_parse!` macro is a monstrosity, and expands into many different identifiers. The hope is that chipping away at things like this will not only make the template easier to understand by framing portions of the problem in terms of more traditional Rust code, but will also hopefully reduce compile times by reducing the amount of code that is expanded by the macro. DEV-7145	2022-09-13 09:20:29 -04:00
Mike Gerwitz	419b24f251	tamer: Introduce NIR (accepting only) This introduces NIR, but only as an accepting grammar; it doesn't yet emit the NIR IR, beyond TODOs. This modifies `tamec` to, while copying XIR, also attempt to lower NIR to produce parser errors, if any. It does not yet fail compilation, as I just want to be cautious and observe that everything's working properly for a little while as people use it, before I potentially break builds. This is the culmination of months of supporting effort. The NIR grammar is derived from our existing TAME sources internally, which I use for now as a test case until I introduce test cases directly into TAMER later on (I'd do it now, if I hadn't spent so much time on this; I'll start introducing tests as I begin emitting NIR tokens). This is capable of fully parsing our largest system with >900 packages, as well as `core`. `tamec`'s lowering is a mess; that'll be cleaned up in future commits. The same can be said about `tameld`. NIR's grammar has some initial documentation, but this will improve over time as well. The generated docs still need some improvement, too, especially with generated identifiers; I just want to get this out here for testing. DEV-7145	2022-08-29 15:52:04 -04:00
Mike Gerwitz	c420ab2730	tamer: xir::parse: Correct doc xrefs These weren't causing problems until they were output as part of NIR (in a separate module). NIR is about to be committed. DEV-7145	2022-08-29 15:52:04 -04:00
Mike Gerwitz	638a9c483b	tamer: xir::parse::ele: Hide internal NT enum variants The user never sees or interacts with these; they're macro-generated, and distract from the useful information in the generated docs. DEV-7145	2022-08-29 15:52:04 -04:00
Mike Gerwitz	2b33a45985	tamer: xir::parse::ele: Support NT docs This just modifies the macro to proxy attributes to generated NTs so that they can be documented. DEV-7145	2022-08-29 15:52:04 -04:00
Mike Gerwitz	51728545f7	tamer: xir::parse::ele: Properly handle previous state transitions This includes when on the last state / expecting a close. Previously, there were a couple major issues: 1. After parsing an NT, we can't allow preemption because we must emit a dead state so that we can remove the NT from the stack, otherwise they'll never close (until the parent does) and that results in unbounded stack growth for a lot of siblings. Therefore, we cannot preempt on `Text`, which causes the NT to receive it, emit a dead state, transition away from the NT, and not accept another NT of the same type after `Text`. 2. When encountering an unknown element, the error message stated that a closing tag was expected rather than one of the elements accepted by the final NT. For #1, this was solved by allowing the parent to transition back to the NT if it would have been matched by the previous NT. A future change may therefore allow us to remove repetition handling entirely and allow the parent to deal with it (maybe). For #2, the trouble is with the parser generator macro---we don't have a good way of knowing the last NT, and the last NT may not even exist if none was provided. This solution is a compromise, after having tried and failed at many others; I desperately need to move on, and this results in the correct behavior and doesn't sacrifice performance. But it can be done better in the future. It's also worth noting for #2 that the behavior isn't _entirely_ desirable, but in practice it is mostly correct. Specifically, if we encounter an unknown token, we're going to blow through all NTs until the last one, which will be forced to handle it. After that, we cannot return to a previous NT, and so we've forefitted the ability to parse anything that came before it. NIR's grammar is such that sequences are rare and, if present, there's really only ever two NTs, and so this awkward behavior will rarely cause practical issues. With that said, it ought to be improved in the future, but let's wait to see if other parts of the lowering pipeline provide more appropriate places to handle some of these things (even though it really ought to be handled at the grammar level). But I'm well out of time to spend on this. I have to move on. DEV-7145	2022-08-29 15:52:04 -04:00
Mike Gerwitz	7466ecbe8b	tamer: xir::parse::ele: Accept missing child `ele_parse!` was recently converted to accept zero-or-more for every NT to simplify the parser-generator, since NIR isn't going to be able to accurately determine whether child requirements are met anyway (because of the template system). This ensures that `Close` can be accepted when we're expecting an element. It also adds a test for a scenario that's causing me some trouble in stashed code so that I can ensure that it doesn't break. DEV-7145	2022-08-22 09:43:59 -04:00
Mike Gerwitz	9366c0c154	tamer: xir::parse::ele: Increase parser nesting depth This sets the maximum depth to 64, which is still arbitrary, but unfortunately the sum types introduce multiple levels of nesting, in particular for template applications, so nested applications can result in a fairly large stack. I have various ideas to improve upon that---limited a bit in that repetition as it is current implemented inhibits tail calls---but they're not worth doing just yet relative to other priorities. The impact of this change is not significant. DEV-7145	2022-08-18 16:16:45 -04:00
Mike Gerwitz	abb2c80e22	tamer: xir::parse::ele: Always repeat This removes support for configurable repetition. What? Why? As it turns out, the complexity that repetition adds is quite significant and is not worth the effort. The truth is that NIR is going to have to allow zero-or-more matches on virtually everything _anyway_ because template application is allowed virtually anywhere---it is not possible to fully statically analyze TAME's sources because templates can expand into just about anything. Given that, AIR (or something down the line) is going to have to supply the necessary invariants instead. It does suck, though, that this removes a lot of code that I fairly recently wrote, and spent a decent amount of time on. But it's important to know when to cut your losses. Perhaps I could have planned better, but deriving this whole system as been quite the experiment. DEV-7145	2022-08-18 15:19:40 -04:00

1 2 3 4

189 Commits (7857460c1dcd74ca6256c836af2d15af73f1c7bc)