cpython

Commit Graph

Author	SHA1	Message	Date
Pablo Galindo	45cf5db587	Allow pgen to produce a DOT format dump of the grammar (GH-18005) Originally suggested by Anthony Shaw.	2020-01-14 22:32:55 +00:00
Vinay Sajip	0b60f64e43	bpo-11410: Standardize and use symbol visibility attributes across POSIX and Windows. (GH-16347)	2019-10-15 08:26:12 +01:00
Pablo Galindo	c638521dbf	Fix typo in the algorithm description (GH-15774)	2019-09-09 15:08:23 +01:00
Shashi Ranjan	43710b67b3	Fix typos in the documentation of Parser/pgen (GH-15416) Co-Authored-By: Antoine <43954001+awecx@users.noreply.github.com>	2019-08-24 19:07:24 +01:00
Pablo Galindo	71876fa438	Refactor Parser/pgen and add documentation and explanations (GH-15373) * Refactor Parser/pgen and add documentation and explanations To improve the readability and maintainability of the parser generator perform the following transformations: * Separate the metagrammar parser in its own class to simplify the parser generator logic. * Create separate classes for DFAs and NFAs and move methods that act exclusively on them from the parser generator to these classes. * Add docstrings and comment documenting the process to go from the grammar file into NFAs and then DFAs. Detail some of the algorithms and give some background explanations of some concepts that will helps readers not familiar with the parser generation process. * Select more descriptive names for some variables and variables. * PEP8 formatting and quote-style homogenization. The output of the parser generator remains the same (Include/graminit.h and Python/graminit.c remain untouched by running the new parser generator).	2019-08-22 02:38:39 +01:00
Hansraj Das	e018dc52d1	Remove duplicate call to strip method in Parser/pgen/token.py (GH-14938)	2019-07-24 21:31:19 +01:00
tyomitch	84b4784f12	use `const` in graminit.c (GH-12713)	2019-04-23 18:29:57 +09:00
Pablo Galindo	f2cf1e3e28	bpo-36623: Clean parser headers and include files (GH-12253) After the removal of pgen, multiple header and function prototypes that lack implementation or are unused are still lying around.	2019-04-13 17:05:14 +01:00
Pablo Galindo	91759d9801	bpo-36143: Regenerate Lib/keyword.py from the Grammar and Tokens file using pgen (GH-12456) Now that the parser generator is written in Python (Parser/pgen) we can make use of it to regenerate the Lib/keyword file that contains the language keywords instead of parsing the autogenerated grammar files. This also allows checking in the CI that the autogenerated files are up to date.	2019-03-25 22:01:12 +00:00
tyomitch	1b304f992d	Remove d_initial from the parser as it is unused (GH-12212) d_initial, the first state of a particular DFA in the parser has always been initialized to 0 in the old pgen as well as the new pgen. As this value is not used and the first state of each DFA is assumed to be the first element in the array representing it, remove d_initial from the parser to reduce complexity.	2019-03-09 15:35:50 +00:00
Pablo Galindo	8bc401a55c	Clean implementation of Parser/pgen and fix some style issues (GH-12156)	2019-03-04 07:26:13 +00:00
Pablo Galindo	1f24a719e7	bpo-35808: Retire pgen and use pgen2 to generate the parser (GH-11814) Pgen is the oldest piece of technology in the CPython repository, building it requires various #if[n]def PGEN hacks in other parts of the code and it also depends more and more on CPython internals. This commit removes the old pgen C code and replaces it for a new version implemented in pure Python. This is a modified and adapted version of lib2to3/pgen2 that can generate grammar files compatibles with the current parser. This commit also eliminates all the #ifdef and code branches related to pgen, simplifying the code and making it more maintainable. The regen-grammar step now uses $(PYTHON_FOR_REGEN) that can be any version of the interpreter, so the new pgen code maintains compatibility with older versions of the interpreter (this also allows regenerating the grammar with the current CI solution that uses Python3.5). The new pgen Python module also makes use of the Grammar/Tokens file that holds the token specification, so is always kept in sync and avoids having to maintain duplicate token definitions.	2019-03-01 15:34:44 -08:00

12 Commits