cpython

Commit Graph

Author	SHA1	Message	Date
Lysandros Nikolaou	1f0f4abb11	bpo-41076: Pre-feed the parser with the f-string expression location (GH-21054) This commit changes the parsing of f-string expressions with the new parser. The parser gets pre-fed with the location of the expression itself (not the f-string, which was what we were doing before). This allows us to completely skip the shifting of the AST nodes after the parsing is completed.	2020-06-28 00:41:48 +01:00
Batuhan Taskaya	c8f29ad986	bpo-40769: Allow extra surrounding parentheses for invalid annotated assignment rule (GH-20387)	2020-06-27 19:33:08 +01:00
Lysandros Nikolaou	6dcbc2422d	bpo-41132: Use pymalloc allocator in the f-string parser (GH-21173)	2020-06-27 18:47:00 +01:00
Lysandros Nikolaou	2e0a920e9e	bpo-41084: Adjust message when an f-string expression causes a SyntaxError (GH-21084) Prefix the error message with `fstring: `, when parsing an f-string expression throws a `SyntaxError`.	2020-06-26 12:24:05 +01:00
Lysandros Nikolaou	4b85e60601	bpo-41119: Output correct error message for list/tuple followed by colon (GH-21160)	2020-06-26 00:22:36 +01:00
Lysandros Nikolaou	564cd18767	bpo-40939: Rename PyPegen* functions to PyParser* (GH-21016) Rename PyPegen* functions to PyParser, so that we can remove the old set of PyParser functions that were using the old parser.	2020-06-22 00:47:46 +01:00
Lysandros Nikolaou	6c4e0bd974	bpo-41060: Avoid SEGFAULT when calling GET_INVALID_TARGET in the grammar (GH-21020) `GET_INVALID_TARGET` might unexpectedly return `NULL`, which if not caught will cause a SEGFAULT. Therefore, this commit introduces a new inline function `RAISE_SYNTAX_ERROR_INVALID_TARGET` that always checks for `GET_INVALID_TARGET` returning NULL and can be used in the grammar, replacing the long C ternary operation used till now.	2020-06-21 03:18:01 +01:00
Lysandros Nikolaou	314858e276	bpo-40939: Remove the old parser (Part 2) (GH-21005) Remove some remaining files and Makefile targets for the old parser	2020-06-20 19:07:25 +01:00
Lysandros Nikolaou	861efc6e8f	bpo-40958: Avoid 'possible loss of data' warning on Windows (GH-20970)	2020-06-20 05:57:27 -07:00
Lysandros Nikolaou	01ece63d42	bpo-40334: Produce better error messages on invalid targets (GH-20106) The following error messages get produced: - `cannot delete ...` for invalid `del` targets - `... is an illegal 'for' target` for invalid targets in for statements - `... is an illegal 'with' target` for invalid targets in with statements Additionally, a few `cut`s were added in various places before the invocation of the `invalid_*` rule, in order to speed things up. Co-authored-by: Pablo Galindo <Pablogsal@gmail.com>	2020-06-19 00:10:43 +01:00
Pablo Galindo	51c5896b62	bpo-40958: Avoid buffer overflow in the parser when indexing the current line (GH-20875)	2020-06-16 16:49:43 +01:00
Pablo Galindo	e0bec69854	Remove old comment in string_parser.c (GH-20906)	2020-06-16 02:13:33 +01:00
Victor Stinner	e822e37946	bpo-36020: Remove snprintf macro in pyerrors.h (GH-20889) On Windows, #include "pyerrors.h" no longer defines "snprintf" and "vsnprintf" macros. PyOS_snprintf() and PyOS_vsnprintf() should be used to get portable behavior. Replace snprintf() calls with PyOS_snprintf() and replace vsnprintf() calls with PyOS_vsnprintf().	2020-06-15 21:59:47 +02:00
Pablo Galindo	fb61c42361	Improve readability and style in parser files (GH-20884)	2020-06-15 14:23:43 +01:00
Pablo Galindo	1ed83adb0e	bpo-40939: Remove the old parser (GH-20768) This commit removes the old parser, the deprecated parser module, the old parser compatibility flags and environment variables and all associated support code and documentation.	2020-06-11 17:30:46 +01:00
Lysandros Nikolaou	bcd7deed91	bpo-40939: Remove PEG parser easter egg (__new_parser__) (#20802 ) It no longer serves a purpose (there's only one parser) and having "new" in any name will eventually look odd. Also, it impinges on a potential sub-namespace, `__new_...__`.	2020-06-11 09:09:21 -07:00
Lysandros Nikolaou	896f4cf63f	bpo-40847: Consider a line with only a LINECONT a blank line (GH-20769) A line with only a line continuation character should be considered a blank line at tokenizer level so that only a single NEWLINE token gets emitted. The old parser was working around the issue, but the new parser threw a `SyntaxError` for valid input. For example, an empty line following a line continuation character was interpreted as a `SyntaxError`. Co-authored-by: Pablo Galindo <Pablogsal@gmail.com>	2020-06-11 00:56:08 +01:00
Victor Stinner	1bcc32f062	bpo-39465: Use _PyInterpreterState_GET() (GH-20788) Replace _PyThreadState_GET() with _PyInterpreterState_GET() in: * get_small_int() * gcmodule.c: add also get_gc_state() function * _PyTrash_deposit_object() * _PyTrash_destroy_chain() * warnings_get_state() * Py_GetRecursionLimit() Cleanup listnode.c: add 'parser' variable.	2020-06-10 20:08:26 +02:00
Pablo Galindo	c6483c9896	Raise specialised syntax error for invalid lambda parameters (GH-20776)	2020-06-10 14:07:06 +01:00
Pablo Galindo	9f495908c5	bpo-40903: Handle multiple '=' in invalid assignment rules in the PEG parser (GH-20697) Automerge-Triggered-By: @pablogsal	2020-06-07 18:57:00 -07:00
Pablo Galindo	972ab03276	bpo-40904: Fix segfault in the new parser with f-string containing yield statements with no value (GH-20701)	2020-06-08 01:47:37 +01:00
Pablo Galindo	2e6593db00	bpo-40880: Fix invalid read in newline_in_string in pegen.c (#20666 ) * bpo-40880: Fix invalid read in newline_in_string in pegen.c * Update Parser/pegen/pegen.c Co-authored-by: Lysandros Nikolaou <lisandrosnik@gmail.com> * Add NEWS entry Co-authored-by: Lysandros Nikolaou <lisandrosnik@gmail.com>	2020-06-06 00:52:27 +01:00
Pablo Galindo	a54096e305	bpo-40883: Fix memory leak in fstring_compile_expr in parse_string.c (GH-20667)	2020-06-06 00:52:15 +01:00
Victor Stinner	fa7ab6aa0f	bpo-40826: Add _PyOS_InterruptOccurred(tstate) function (GH-20599) my_fgets() now calls _PyOS_InterruptOccurred(tstate) to check for pending signals, rather calling PyOS_InterruptOccurred(). my_fgets() is called with the GIL released, whereas PyOS_InterruptOccurred() must be called with the GIL held. test_repl: use text=True and avoid SuppressCrashReport in test_multiline_string_parsing(). Fix my_fgets() on Windows: fgets(fp) does crash if fileno(fp) is closed.	2020-06-03 14:39:59 +02:00
Victor Stinner	c353764fd5	bpo-40826: Fix GIL usage in PyOS_Readline() (GH-20579) Fix GIL usage in PyOS_Readline(): lock the GIL to set an exception. Pass tstate to my_fgets() and _PyOS_WindowsConsoleReadline(). Cleanup these functions.	2020-06-01 20:59:35 +02:00
Shantanu	c116c94ff1	bpo-40614: Respect feature version for f-string debug expressions (GH-20196) Co-authored-by: Lysandros Nikolaou <lisandrosnik@gmail.com> Co-authored-by: Pablo Galindo <pablogsal@gmail.com>	2020-05-27 21:30:38 +01:00
Lysandros Nikolaou	526e23f153	Refactor error handling code in Parser/pegen/pegen.c (GH-20440) Set p->error_indicator in various places, where it's needed, but it's not done. Automerge-Triggered-By: @gvanrossum	2020-05-27 09:04:11 -07:00
Pablo Galindo	1cf15af9a6	bpo-40217: Ensure Py_VISIT(Py_TYPE(self)) is always called for PyType_FromSpec types (reverts GH-19414) (GH-20264) Heap types now always visit the type in tp_traverse. See added docs for details. This reverts commit `0169d3003b`. Automerge-Triggered-By: @encukou	2020-05-27 02:03:38 -07:00
Pablo Galindo	404b23b85b	Fix lookahead of soft keywords in the PEG parser (GH-20436) Automerge-Triggered-By: @gvanrossum	2020-05-26 16:15:52 -07:00
Guido van Rossum	b45af1a569	Add soft keywords (GH-20370) These are like keywords but they only work in context; they are not reserved except when there is an exact match. This would enable things like match statements without reserving `match` (which would be bad for the `re.match()` function and probably lots of other places). Automerge-Triggered-By: @gvanrossum	2020-05-26 10:58:44 -07:00
Ammar Askar	a2bbedc8b1	Fix peg_generator compiler warnings under MSVC (GH-20405)	2020-05-26 05:33:35 +01:00
Lysandros Nikolaou	f7b1e46156	bpo-38964: Print correct filename on a SyntaxError in an fstring (GH-20399) When a `SyntaxError` in the expression part of a fstring is found, the filename attribute of the `SyntaxError` is always `<fstring>`. With this commit, it gets changed to always have the name of the file the fstring resides in. Co-authored-by: Pablo Galindo <Pablogsal@gmail.com>	2020-05-26 01:32:18 +01:00
Pablo Galindo	deb4355a37	bpo-40750: Do not expand the new parser debug flags if Py_BUILD_CORE is not defined (GH-20393)	2020-05-25 20:17:12 +01:00
Pablo Galindo	800a35c623	bpo-40750: Support -d flag in the new parser (GH-20340)	2020-05-25 18:38:45 +01:00
Rémi Lapeyre	c73914a562	bpo-36290: Fix keytword collision handling in AST node constructors (GH-12382)	2020-05-24 22:12:57 +01:00
Pablo Galindo	b23d7adfdf	Use Py_ssize_t for the column number in the PEG support code (GH-20341)	2020-05-24 06:01:34 +01:00
Lysandros Nikolaou	ae14583302	bpo-40334: Produce better error messages for non-parenthesized genexps (GH-20153) The error message, generated for a non-parenthesized generator expression in function calls, was still the generic `invalid syntax`, when the generator expression wasn't appearing as the first argument in the call. With this patch, even on input like `f(a, b, c for c in d, e)`, the correct error message gets produced.	2020-05-22 01:56:52 +01:00
Batuhan Taskaya	b8a65ec1d3	bpo-40715: Reject dict unpacking on dict comprehensions (GH-20292) Co-authored-by: Lysandros Nikolaou <lisandrosnik@gmail.com> Co-authored-by: Pablo Galindo <pablogsal@gmail.com>	2020-05-21 23:39:56 +01:00
Batuhan Taskaya	72e0aa2fd2	bpo-40176: Improve error messages for trailing comma on from import (GH-20294)	2020-05-21 21:41:58 +01:00
Pablo Galindo	ced4e5c227	Regenerate the parser (#20195 )	2020-05-18 23:47:51 +02:00
Lysandros Nikolaou	75b863aa97	bpo-40334: Reproduce error message for type comments on bare '*' in the new parser (GH-20151)	2020-05-18 20:14:47 +01:00
Batuhan Taskaya	63b8e0cba3	bpo-40528: Improve AST generation script to do builds simultaneously (GH-19968) - Switch from getopt to argparse. - Removed the limitation of not being able to produce both C and H simultaneously. This will make it run faster since it parses the asdl definition once and uses the generated tree to generate both the header and the C source.	2020-05-18 18:42:10 +01:00
Lysandros Nikolaou	7b7a21bc4f	bpo-40661: Fix segfault when parsing invalid input (GH-20165) Fix segfaults when parsing very complex invalid input, like `import äˆ ð£„¯ð¢·žð±‹á”€ð””ð‘©±å®ä±¬ð©¾\nð—¶½`. Co-authored-by: Guido van Rossum <guido@python.org> Co-authored-by: Pablo Galindo <pablogsal@gmail.com>	2020-05-18 18:32:03 +01:00
Lysandros Nikolaou	2c8cd06afe	bpo-40334: Improvements to error-handling code in the PEG parser (GH-20003) The following improvements are implemented in this commit: - `p->error_indicator` is set, in case malloc or realloc fail. - Avoid memory leaks in the case that realloc fails. - Call `PyErr_NoMemory()` instead of `PyErr_Format()`, because it requires no memory. Co-authored-by: Pablo Galindo <Pablogsal@gmail.com>	2020-05-17 04:19:23 +01:00
Pablo Galindo	16ab07063c	bpo-40334: Correctly identify invalid target in assignment errors (GH-20076) Co-authored-by: Lysandros Nikolaou <lisandrosnik@gmail.com>	2020-05-15 02:04:52 +01:00
Lysandros Nikolaou	ce21cfca7b	bpo-40618: Disallow invalid targets in augassign and except clauses (GH-20083) This commit fixes the new parser to disallow invalid targets in the following scenarios: - Augmented assignments must only accept a single target (Name, Attribute or Subscript), but no tuples or lists. - `except` clauses should only accept a single `Name` as a target. Co-authored-by: Pablo Galindo <Pablogsal@gmail.com>	2020-05-14 21:13:50 +01:00
Pablo Galindo	bcc3036095	bpo-40619: Correctly handle error lines in programs without file mode (GH-20090)	2020-05-14 21:11:48 +01:00
Lysandros Nikolaou	a15c9b3a05	bpo-40334: Always show the caret on SyntaxErrors (GH-20050) This commit fixes SyntaxError locations when the caret is not displayed, by doing the following: - `col_number` always gets set to the location of the offending node/expr. When no caret is to be displayed, this gets achieved by setting the object holding the error line to None. - Introduce a new function `_PyPegen_raise_error_known_location`, which can be called, when an arbitrary `lineno`/`col_offset` needs to be passed. This function then gets used in the grammar (through some new macros and inline functions) so that SyntaxError locations of the new parser match that of the old.	2020-05-13 20:36:27 +01:00
Serhiy Storchaka	74ea6b5a75	bpo-40593: Improve syntax errors for invalid characters in source code. (GH-20033)	2020-05-12 12:42:04 +03:00
Shantanu	27c0d9b54a	bpo-40334: produce specialized errors for invalid del targets (GH-19911)	2020-05-11 14:53:58 -07:00
Pablo Galindo	5b956ca42d	bpo-40585: Normalize errors messages in codeop when comparing them (GH-20030) With the new parser, the error message contains always the trailing newlines, causing the comparison of the repr of the error messages in codeop to fail. This commit makes the new parser mirror the old parser's behaviour regarding trailing newlines.	2020-05-11 01:41:26 +01:00
Pablo Galindo	ac7a92cc0a	bpo-40334: Avoid collisions between parser variables and grammar variables (GH-19987) This is for the C generator: - Disallow rule and variable names starting with `_` - Rename most local variable names generated by the parser to start with `_` Exceptions: - Renaming `p` to `_p` will be a separate PR - There are still some names that might clash, e.g. - anything starting with `Py` - C reserved words (`if` etc.) - Macros like `EXTRA` and `CHECK`	2020-05-09 21:34:50 -07:00
Joannah Nanjekye	d10091aa17	bpo-40502: Initialize n->n_col_offset (GH-19988) * initialize n->n_col_offset * 📜🤖 Added by blurb_it. * Move initialization Co-authored-by: nanjekyejoannah <joannah.nanjekye@ibm.com> Co-authored-by: blurb-it[bot] <43283697+blurb-it[bot]@users.noreply.github.com>	2020-05-08 17:58:28 -03:00
Pablo Galindo	db9163ceef	bpo-40555: Check for p->error_indicator in loop rules after the main loop is done (GH-19986)	2020-05-08 03:38:44 +01:00
Lysandros Nikolaou	4638c64295	bpo-40334: Error message for invalid default args in function call (GH-19973) When parsing something like `f(g()=2)`, where the name of a default arg is not a NAME, but an arbitrary expression, a specialised error message is emitted.	2020-05-07 11:44:06 +01:00
Lysandros Nikolaou	2f37c355ab	bpo-40334: Fix error location upon parsing an invalid string literal (GH-19962) When parsing a string with an invalid escape, the old parser used to point to the beginning of the invalid string. This commit changes the new parser to match that behaviour, since it's currently pointing to the end of the string (or to be more precise, to the beginning of the next token).	2020-05-07 11:37:51 +01:00
Pablo Galindo	470aac4d8e	bpo-40334: Generate comments in the parser code to improve debugging (GH-19966)	2020-05-06 23:14:43 +01:00
Pablo Galindo	99db2a1db7	bpo-40334: Allow trailing comma in parenthesised context managers (GH-19964)	2020-05-06 22:54:34 +01:00
Lysandros Nikolaou	999ec9ab6a	bpo-40334: Add type to the assignment rule in the grammar file (GH-19963)	2020-05-06 19:11:04 +01:00
Batuhan Taskaya	091951a67c	bpo-40528: Improve and clear several aspects of the ASDL definition code for the AST (GH-19952)	2020-05-06 15:29:32 +01:00
Lysandros Nikolaou	846d8b28ab	bpo-40246: Revert reporting of invalid string prefixes (GH-19888) Due to backwards compatibility concerns regarding keywords immediately followed by a string without whitespace between them (like in `bg="#d00" if clear else"#fca"`) will fail to parse, commit `41d5b94af4` has to be reverted.	2020-05-04 12:32:18 +01:00
Lysandros Nikolaou	e10e7c771b	bpo-40334: Spacialized error message for invalid args after bare '' (GH-19865) When parsing things like `def f(): pass` the old parser used to output `SyntaxError: named arguments must follow bare *`, which the new parser wasn't able to do.	2020-05-04 11:58:31 +01:00
Shantanu	c3f001461d	bpo-40491: Fix typo in syntax error for numeric literals (GH-19893)	2020-05-04 11:13:30 +03:00
Shantanu	603d354626	bpo-40493: fix function type comment parsing (GH-19894) The grammar for func_type_input rejected things like `(*t1) ->t2`. This fixes that. Automerge-Triggered-By: @gvanrossum	2020-05-03 22:08:14 -07:00
Lysandros Nikolaou	7f06af684a	bpo-40334: Set error_indicator in _PyPegen_raise_error (GH-19887) Due to PyErr_Occurred not being called at the beginning of each rule, we need to set the error indicator, so that rules do not get expanded after an exception has been thrown	2020-05-04 01:20:09 +01:00
Lysandros Nikolaou	03b7642265	bpo-40334: Make the PyPegen* and PyParser* APIs more consistent (GH-19839) This commit makes both APIs more consistent by doing the following: - Remove the `PyPegen_CodeObjectFrom` functions, which weren't used and will probably not be needed. Functions like `Py_CompileStringObject` can be used instead. - Include a `const char filename` parameter in `PyPegen_ASTFromString`. - Rename `PyPegen_ASTFromFile` to `PyPegen_ASTFromFilename`, because its signature is not the same with `PyParser_ASTFromFile`.	2020-05-01 18:30:51 +01:00
Guido van Rossum	d9d6eadf00	Ensure that tok->type_comments is set on every path (GH-19828)	2020-05-01 17:42:32 +01:00
Guido van Rossum	3941d9700b	bpo-40334: Refactor lambda_parameters similar to parameters (GH-19830)	2020-05-01 17:42:03 +01:00
Pablo Galindo	d955241469	bpo-40334: Correct return value of func_type_comment (GH-19833)	2020-05-01 08:32:09 -07:00
Batuhan Taskaya	76c1b4d5c5	bpo-40334: Improve column offsets for thrown syntax errors by Pegen (GH-19782)	2020-05-01 14:13:43 +01:00
Pablo Galindo	b796b3fb48	bpo-40334: Simplify type handling in the PEG c_generator (GH-19818)	2020-05-01 12:32:26 +01:00
Lysandros Nikolaou	3e0a6f37df	bpo-40334: Add support for feature_version in new PEG parser (GH-19827) `ast.parse` and `compile` support a `feature_version` parameter that tells the parser to parse the input string, as if it were written in an older Python version. The `feature_version` is propagated to the tokenizer, which uses it to handle the three different stages of support for `async` and `await`. Additionally, it disallows the following at parser level: - The '@' operator in < 3.5 - Async functions in < 3.5 - Async comprehensions in < 3.6 - Underscores in numeric literals in < 3.6 - Await expression in < 3.5 - Variable annotations in < 3.6 - Async for-loops in < 3.5 - Async with-statements in < 3.5 - F-strings in < 3.6 Closes we-like-parsers/cpython#124.	2020-04-30 20:27:52 -07:00
Guido van Rossum	c001c09e90	bpo-40334: Support type comments (GH-19780) This implements full support for # type: <type> comments, # type: ignore <stuff> comments, and the func_type parsing mode for ast.parse() and compile(). Closes https://github.com/we-like-parsers/cpython/issues/95. (For now, you need to use the master branch of mypy, since another issue unique to 3.9 had to be fixed there, and there's no mypy release yet.) The only thing missing is `feature_version=N`, which is being tracked in https://github.com/we-like-parsers/cpython/issues/124.	2020-04-30 12:12:19 -07:00
Pablo Galindo	4db245ee9d	bpo-40334: refactor and cleanup for the PEG generators (GH-19775)	2020-04-29 10:42:21 +01:00
Lysandros Nikolaou	6d65087655	bpo-40334: Disallow invalid single statements in the new parser (GH-19774) After parsing is done in single statement mode, the tokenizer buffer has to be checked for additional lines and a `SyntaxError` must be raised, in case there are any. Co-authored-by: Pablo Galindo <Pablogsal@gmail.com>	2020-04-29 02:42:27 +01:00
Pablo Galindo	2208134918	bpo-40334: Explicitly cast to int in pegen.c to fix a compiler warning (GH-19779)	2020-04-29 02:04:06 +01:00
Lysandros Nikolaou	37af21b667	bpo-40334: Fix shifting of nested f-strings in the new parser (GH-19771) `JoinedStr`s and `FormattedValue also needs to be shifted, in order to correctly compute the location information of nested f-strings.	2020-04-29 01:43:50 +01:00
Lysandros Nikolaou	d55133f49f	bpo-40334: Catch E_EOF error, when the tokenizer returns ERRORTOKEN (GH-19743) An E_EOF error was only being caught after the parser exited before this commit. There are some cases though, where the tokenizer returns ERRORTOKEN and has set an E_EOF error (like when EOF directly follows a line continuation character) which weren't correctly handled before.	2020-04-28 01:23:35 +01:00
Pablo Galindo	b94dbd7ac3	bpo-40334: Support PyPARSE_DONT_IMPLY_DEDENT in the new parser (GH-19736)	2020-04-27 18:35:58 +01:00
Pablo Galindo	2b74c835a7	bpo-40334: Support CO_FUTURE_BARRY_AS_BDFL in the new parser (GH-19721) This commit also allows to pass flags to the new parser in all interfaces and fixes a bug in the parser generator that was causing to inline rules with actions, making them disappear.	2020-04-27 18:02:07 +01:00
Pablo Galindo	9f27dd3e16	Use Py_ssize_t instead of ssize_t (GH-19685)	2020-04-24 01:13:33 +01:00
Lysandros Nikolaou	ebebb6429c	bpo-40334: Improve various PEG-Parser related stuff (GH-19669) The changes in this commit are all related to @vstinner's original review comments of the initial PEP 617 implementation PR.	2020-04-23 16:36:06 +01:00
Pablo Galindo	1df5a9e88c	bpo-40334: Fix build errors and warnings in test_peg_generator (GH-19672)	2020-04-23 12:42:13 +01:00
Pablo Galindo	ee40e4b856	bpo-40334: Don't downcast from Py_ssize_t to int (GH-19671)	2020-04-23 03:43:08 +01:00
Pablo Galindo	0b7829e089	Compile extensions in test_peg_generator with C99 (GH-19668)	2020-04-23 03:24:25 +01:00
Pablo Galindo	458004bf79	bpo-40334: Fix errors in parse_string.c with old compilers (GH-19666)	2020-04-23 00:13:47 +01:00
Pablo Galindo	c5fc156852	bpo-40334: PEP 617 implementation: New PEG parser for CPython (GH-19503) Co-authored-by: Guido van Rossum <guido@python.org> Co-authored-by: Lysandros Nikolaou <lisandrosnik@gmail.com>	2020-04-22 23:29:27 +01:00
Pablo Galindo	11a7f158ef	bpo-40335: Correctly handle multi-line strings in tokenize error scenarios (GH-19619) Co-authored-by: Guido van Rossum <gvanrossum@gmail.com>	2020-04-21 01:53:04 +01:00
Lysandros Nikolaou	9a4b38f66b	bpo-40267: Fix message when last input character produces a SyntaxError (GH-19521) When there is a SyntaxError after reading the last input character from the tokenizer and if no newline follows it, the error message used to be `unexpected EOF while parsing`, which is wrong.	2020-04-15 11:22:10 -07:00
Victor Stinner	4a21e57fe5	bpo-40268: Remove unused structmember.h includes (GH-19530) If only offsetof() is needed: include stddef.h instead. When structmember.h is used, add a comment explaining that PyMemberDef is used.	2020-04-15 02:35:41 +02:00
Victor Stinner	62183b8d6d	bpo-40268: Remove explicit pythread.h includes (#19529 ) Remove explicit pythread.h includes: it is always included by Python.h.	2020-04-15 02:04:42 +02:00
Victor Stinner	e5014be049	bpo-40268: Remove a few pycore_pystate.h includes (GH-19510)	2020-04-14 17:52:15 +02:00
Victor Stinner	81a7be3fa2	bpo-40268: Rename _PyInterpreterState_GET_UNSAFE() (GH-19509) Rename _PyInterpreterState_GET_UNSAFE() to _PyInterpreterState_GET() for consistency with _PyThreadState_GET() and to have a shorter name (help to fit into 80 columns). Add also "assert(tstate != NULL);" to the function.	2020-04-14 15:14:01 +02:00
Victor Stinner	4a3fe08353	bpo-40268: Include explicitly pycore_interp.h (GH-19505) pycore_pystate.h no longer includes pycore_interp.h: it's now included explicitly in files accessing PyInterpreterState.	2020-04-14 14:26:24 +02:00
Lysandros Nikolaou	41d5b94af4	bpo-40246: Report a better error message for invalid string prefixes (GH-19476)	2020-04-12 19:21:00 +01:00
Pablo Galindo	168660b547	bpo-40141: Add line and column information to ast.keyword nodes (GH-19283)	2020-04-02 00:47:39 +01:00
Alexander Riccio	51e3e450fb	bpo-40020: Fix realloc leak on failure in growable_comment_array_add (GH-19083) Fix a leak and subsequent crash in parsetok.c caused by realloc misuse on a rare codepath. Realloc returns a null pointer on failure, and then growable_comment_array_deallocate crashes later when it dereferences it.	2020-03-30 23:15:59 +02:00
Victor Stinner	87d3b9db4a	bpo-39882: Add _Py_FatalErrorFormat() function (GH-19157)	2020-03-25 19:27:36 +01:00
Serhiy Storchaka	bace59d8b8	bpo-39999: Improve compatibility of the ast module. (GH-19056) * Re-add removed classes Suite, slice, Param, AugLoad and AugStore. * Add docstrings for dummy classes. * Add docstrings for attribute aliases. * Set __module__ to "ast" instead of "_ast".	2020-03-22 20:33:34 +02:00
Serhiy Storchaka	6b97598fb6	bpo-39988: Remove ast.AugLoad and ast.AugStore node classes. (GH-19038)	2020-03-17 23:41:08 +02:00
Batuhan Taşkaya	4ab362cec6	bpo-39638: Keep ASDL signatures in the AST nodes (GH-18515)	2020-03-16 10:12:53 +02:00
Batuhan Taşkaya	8689209e03	bpo-39969: Remove ast.Param node class as is no longer used (GH-19020)	2020-03-15 19:32:17 +00:00
Serhiy Storchaka	13d52c2686	bpo-34822: Simplify AST for subscription. (GH-9605) * Remove the slice type. * Make Slice a kind of the expr type instead of the slice type. * Replace ExtSlice(slices) with Tuple(slices, Load()). * Replace Index(value) with a value itself. All non-terminal nodes in AST for expressions are now of the expr type.	2020-03-10 18:52:34 +02:00
Serhiy Storchaka	b7e9525f9c	bpo-36287: Make ast.dump() not output optional fields and attributes with default values. (GH-18843) The default values for optional fields and attributes of AST nodes are now set as class attributes (e.g. Constant.kind is set to None).	2020-03-10 00:07:47 +02:00
xatier	d7a04a8425	Fix typo in the parser generator (GH-18603)	2020-03-09 02:58:24 +00:00
Victor Stinner	9e5d30cc99	bpo-39882: Py_FatalError() logs the function name (GH-18819) The Py_FatalError() function is replaced with a macro which logs automatically the name of the current function, unless the Py_LIMITED_API macro is defined. Changes: * Add _Py_FatalErrorFunc() function. * Remove the function name from the message of Py_FatalError() calls which included the function name. * Update tests.	2020-03-07 00:54:20 +01:00
Batuhan Taşkaya	d82e469048	bpo-39639: Remove the AST "Suite" node and associated code (GH-18513) The AST "Suite" node is no longer used and it can be removed from the ASDL definition and related structures (compiler, visitors, ...). Co-Authored-By: Victor Stinner <vstinner@python.org> Co-authored-by: Brett Cannon <54418+brettcannon@users.noreply.github.com> Co-authored-by: Pablo Galindo <Pablogsal@gmail.com>	2020-03-04 16:16:46 +00:00
Andy Lester	384f3c536d	closes bpo-39721: Fix constness of members of tok_state struct. (GH-18600) The function PyTokenizer_FromUTF8 from Parser/tokenizer.c had a comment: /* XXX: constify members. / This patch addresses that. In the tok_state struct: end and start were non-const but could be made const * str and input were const but should have been non-const Changes to support this include: * decode_str() now returns a char * since it is allocated. * PyTokenizer_FromString() and PyTokenizer_FromUTF8() each creates a new char * for an allocate string instead of reusing the input const char . PyTokenizer_Get() and tok_get() now take const char ** arguments. * Various local vars are const or non-const accordingly. I was able to remove five casts that cast away constness.	2020-02-27 18:44:52 -08:00
Serhiy Storchaka	0cc6b5e559	bpo-39219: Fix SyntaxError attributes in the tokenizer. (GH-17828) * Always set the text attribute. * Correct the offset attribute for non-ascii sources.	2020-02-12 12:17:00 +02:00
Victor Stinner	f3e7ea5b8c	bpo-39500: Document PyUnicode_IsIdentifier() function (GH-18397) PyUnicode_IsIdentifier() does not call Py_FatalError() anymore if the string is not ready.	2020-02-11 14:29:33 +01:00
Brandt Bucher	d2f9667264	bpo-38823: Fix refleaks in _ast initialization error path (GH-17276)	2020-02-06 15:45:46 +01:00
Pablo Galindo	45cf5db587	Allow pgen to produce a DOT format dump of the grammar (GH-18005) Originally suggested by Anthony Shaw.	2020-01-14 22:32:55 +00:00
Emmanuel Arias	d23f78267a	Remove unused functions in Parser/parsetok.c (GH-17365)	2020-01-13 11:58:52 +00:00
Alex Henrie	7ba6f18de2	bpo-39307: Fix memory leak on error path in parsetok (GH-17953)	2020-01-13 10:35:47 +00:00
Pablo Galindo	5ec91f78d5	bpo-39209: Manage correctly multi-line tokens in interactive mode (GH-17860)	2020-01-06 15:59:09 +00:00
Steve Dower	a9d0a6a1b9	bpo-36500: Simplify PCbuild/build.bat and prevent path separator changing in comments (GH-17644)	2019-12-17 14:14:13 -08:00
Batuhan Taşkaya	109fc2792a	bpo-38673: dont switch to ps2 if the line starts with comment or whitespace (GH-17421) https://bugs.python.org/issue38673	2019-12-08 20:36:27 -08:00
Vinay Sajip	9def81aa52	bpo-36876: Moved Parser/listnode.c statics to interpreter state. (GH-16328)	2019-11-07 10:08:58 +00:00
Max Bernstein	bdac32e9fe	closes bpo-38648: Remove double tp_free slot in Python-ast.c. (GH-17002) This looks like a typo due to copy-paste.	2019-10-30 18:08:06 -07:00
Vinay Sajip	0b60f64e43	bpo-11410: Standardize and use symbol visibility attributes across POSIX and Windows. (GH-16347)	2019-10-15 08:26:12 +01:00
Dong-hee Na	a05fcd3c7a	bpo-38425: Fix ‘res’ may be used uninitialized warning (GH-16688)	2019-10-10 09:41:26 +02:00
Eddie Elizondo	3368f3c6ae	bpo-38140: Make dict and weakref offsets opaque for C heap types (#16076 ) * Make dict and weakref offsets opaque for C heap types * Add news	2019-09-19 17:29:05 +01:00
Eddie Elizondo	0247e80f3c	Fix leaks in Python-ast.c (#16127 )	2019-09-14 14:38:17 +01:00
Zackery Spytz	421a72af4d	bpo-21120: Exclude Python-ast.h, ast.h and asdl.h from the limited API (#14634 ) The PyArena type is not part of the limited API, so these headers shouldn't be part of it either.	2019-09-12 10:27:14 +01:00
Dino Viehland	ac46eb4ad6	bpo-38113: Update the Python-ast.c generator to PEP384 (gh-15957) Summary: This mostly migrates Python-ast.c to PEP384 and removes all statics from the whole file. This modifies the generator itself that generates the Python-ast.c. It leaves in the usage of _PyObject_LookupAttr even though it's not fully PEP384 compatible (this could always be shimmed in by anyone who needs it).	2019-09-11 18:16:34 +01:00
Serhiy Storchaka	43c9731334	bpo-38083: Minor improvements in asdl_c.py and Python-ast.c. (GH-15824) * Use the const qualifier for constant C strings. * Intern field and attribute names. * Temporary incref a borrowed reference to a list item.	2019-09-10 03:02:30 -07:00
Greg Price	fa3a38d81f	Mark files as executable that are meant as scripts. (GH-15354) This is the converse of GH-15353 -- in addition to plenty of scripts in the tree that are marked with the executable bit (and so can be directly executed), there are a few that have a leading `#!` which could let them be executed, but it doesn't do anything because they don't have the executable bit set. Here's a command which finds such files and marks them. The first line finds files in the tree with a `#!` line anywhere; the next-to-last step checks that the first line is actually of that form. In between we filter out files that already have the bit set, and some files that are meant as fragments to be consumed by one or another kind of preprocessor. $ git grep -l '^#!' \ \| grep -vxFf <( \ git ls-files --stage \ \| perl -lane 'print $F[3] if (!/^100644/)' \ ) \ \| grep -ve '\.in$' -e '^Doc/includes/' \ \| while read f; do head -c2 "$f" \| grep -qxF '#!' \ && chmod a+x "$f"; \ done	2019-09-09 07:16:33 -07:00
Pablo Galindo	c638521dbf	Fix typo in the algorithm description (GH-15774)	2019-09-09 15:08:23 +01:00
Shashi Ranjan	43710b67b3	Fix typos in the documentation of Parser/pgen (GH-15416) Co-Authored-By: Antoine <43954001+awecx@users.noreply.github.com>	2019-08-24 19:07:24 +01:00
Pablo Galindo	71876fa438	Refactor Parser/pgen and add documentation and explanations (GH-15373) * Refactor Parser/pgen and add documentation and explanations To improve the readability and maintainability of the parser generator perform the following transformations: * Separate the metagrammar parser in its own class to simplify the parser generator logic. * Create separate classes for DFAs and NFAs and move methods that act exclusively on them from the parser generator to these classes. * Add docstrings and comment documenting the process to go from the grammar file into NFAs and then DFAs. Detail some of the algorithms and give some background explanations of some concepts that will helps readers not familiar with the parser generation process. * Select more descriptive names for some variables and variables. * PEP8 formatting and quote-style homogenization. The output of the parser generator remains the same (Include/graminit.h and Python/graminit.c remain untouched by running the new parser generator).	2019-08-22 02:38:39 +01:00
Hansraj Das	69f37bcb28	Indent code inside if block. (GH-15284) Without indendation, seems like strcpy line is parallel to `if` condition.	2019-08-15 09:19:07 -07:00
Anthony Sottile	5b94f3578c	Fix `SyntaxError` indicator printing too many spaces for multi-line strings (GH-14433)	2019-07-29 14:59:13 +01:00
Hansraj Das	e018dc52d1	Remove duplicate call to strip method in Parser/pgen/token.py (GH-14938)	2019-07-24 21:31:19 +01:00
Pablo Galindo	cd6e83b481	bpo-37593: Swap the positions of posonlyargs and args in the constructor of ast.parameters nodes (GH-14778) https://bugs.python.org/issue37593	2019-07-14 16:32:18 -07:00
Victor Stinner	022ac0a497	bpo-37253: Remove PyAST_obj2mod_ex() function (GH-14020) PyAST_obj2mod_ex() is similar to PyAST_obj2mod() with an additional 'feature_version' parameter which is unused.	2019-06-13 09:18:45 +02:00
Jeroen Demeyer	530f506ac9	bpo-36974: tp_print -> tp_vectorcall_offset and tp_reserved -> tp_as_async (GH-13464) Automatically replace tp_print -> tp_vectorcall_offset tp_compare -> tp_as_async tp_reserved -> tp_as_async	2019-05-30 19:13:39 -07:00
Eric V. Smith	6f6ff8a565	bpo-37050: Remove expr_text from FormattedValue ast node, use Constant node instead (GH-13597) When using the "=" debug functionality of f-strings, use another Constant node (or a merged constant node) instead of adding expr_text to the FormattedValue node.	2019-05-27 15:31:52 -04:00
Steve Dower	b82e17e626	bpo-36842: Implement PEP 578 (GH-12613) Adds sys.audit, sys.addaudithook, io.open_code, and associated C APIs.	2019-05-23 08:45:22 -07:00
Michael J. Sullivan	d8a82e2897	bpo-36878: Only allow text after `# type: ignore` if first character ASCII (GH-13504) This disallows things like `# type: ignoreé`, which seems wrong. Also switch to using Py_ISALNUM for the alnum check, for consistency with other code (and maybe correctness re: locale issues?). https://bugs.python.org/issue36878	2019-05-22 13:43:36 -07:00
Michael J. Sullivan	933e1509ec	bpo-36878: Track extra text added to 'type: ignore' in the AST (GH-13479) GH-13238 made extra text after a # type: ignore accepted by the parser. This finishes the job and actually plumbs the extra text through the parser and makes it available in the AST.	2019-05-22 15:54:20 +01:00
Matthias Bussonnier	565b4f1ac7	bpo-34616: Add PyCF_ALLOW_TOP_LEVEL_AWAIT to allow top-level await (GH-13148) Co-Authored-By: Yury Selivanov <yury@magic.io>	2019-05-21 16:12:02 -04:00
Anthony Sottile	abea73bf4a	bpo-2180: Treat line continuation at EOF as a `SyntaxError` (GH-13401) This makes the parser consistent with the tokenize module (already the case in `pypy`). sample ------ ```python x = 5\ ``` before ------ ```console $ python3 t.py $ python3 -mtokenize t.py t.py:2:0: error: EOF in multi-line statement ``` after ----- ```console $ ./python t.py File "t.py", line 3 x = 5\ ^ SyntaxError: unexpected EOF while parsing $ ./python -m tokenize t.py t.py:2:0: error: EOF in multi-line statement ``` https://bugs.python.org/issue2180	2019-05-18 11:27:16 -07:00
Michael J. Sullivan	d8320ecb86	bpo-36878: Allow extra text after `# type: ignore` comments (GH-13238) In the parser, when using the type_comments=True option, recognize a TYPE_IGNORE as anything containing `# type: ignore` followed by a non-alphanumeric character. This is to allow ignores such as `# type: ignore[E1000]`.	2019-05-11 19:17:24 +01:00
Eric V. Smith	9a4135e939	bpo-36817: Add f-string debugging using '='. (GH-13123) If a "=" is specified a the end of an f-string expression, the f-string will evaluate to the text of the expression, followed by '=', followed by the repr of the value of the expression.	2019-05-08 16:28:48 -04:00
Pablo Galindo	8c77b8cb91	bpo-36540: PEP 570 -- Implementation (GH-12701) This commit contains the implementation of PEP570: Python positional-only parameters. * Update Grammar/Grammar with new typedarglist and varargslist * Regenerate grammar files * Update and regenerate AST related files * Update code object * Update marshal.c * Update compiler and symtable * Regenerate importlib files * Update callable objects * Implement positional-only args logic in ceval.c * Regenerate frozen data * Update standard library to account for positional-only args * Add test file for positional-only args * Update other test files to account for positional-only args * Add News entry * Update inspect module and related tests	2019-04-29 13:36:57 +01:00
Inada Naoki	09415ff0eb	fix warnings by adding more const (GH-12924)	2019-04-23 20:39:37 +09:00
tyomitch	84b4784f12	use `const` in graminit.c (GH-12713)	2019-04-23 18:29:57 +09:00
Pablo Galindo	f2cf1e3e28	bpo-36623: Clean parser headers and include files (GH-12253) After the removal of pgen, multiple header and function prototypes that lack implementation or are unused are still lying around.	2019-04-13 17:05:14 +01:00
Zackery Spytz	cda139d1de	bpo-36459: Fix a possible double PyMem_FREE() due to tokenizer.c's tok_nextc() (12601) Remove the PyMem_FREE() call added in `cb90c89`. The buffer will be freed when PyTokenizer_Free() is called on the tokenizer state.	2019-03-28 15:53:00 +02:00
Pablo Galindo	91759d9801	bpo-36143: Regenerate Lib/keyword.py from the Grammar and Tokens file using pgen (GH-12456) Now that the parser generator is written in Python (Parser/pgen) we can make use of it to regenerate the Lib/keyword file that contains the language keywords instead of parsing the autogenerated grammar files. This also allows checking in the CI that the autogenerated files are up to date.	2019-03-25 22:01:12 +00:00
Emmanuel Arias	ed5e29cba5	bpo-36385: Add ``elif`` sentence on to avoid multiple ``if`` (GH-12478) Currently, when arguments on Parser/asdl_c.py are parsed ``ìf`` sentence is used. This PR Propose to use ``elif`` to avoid multiple evaluting of the ifs. https://bugs.python.org/issue36385	2019-03-20 21:39:17 -07:00
Pablo Galindo	cb90c89de1	bpo-36367: Free buffer if realloc fails in tokenize.c (GH-12442)	2019-03-19 17:17:58 +00:00
Guido van Rossum	10f8ce6688	bpo-36280: Add Constant.kind field (GH-12295) The value is a string for string and byte literals, None otherwise. It is 'u' for u"..." literals, 'b' for b"..." literals, '' for "..." literals. The 'r' (raw) prefix is ignored. Does not apply to f-strings. This appears sufficient to make mypy capable of using the stdlib ast module instead of typed_ast (assuming a mypy patch I'm working on). WIP: I need to make the tests pass. @ilevkivskyi @serhiy-storchaka https://bugs.python.org/issue36280	2019-03-13 13:00:46 -07:00
tyomitch	1b304f992d	Remove d_initial from the parser as it is unused (GH-12212) d_initial, the first state of a particular DFA in the parser has always been initialized to 0 in the old pgen as well as the new pgen. As this value is not used and the first state of each DFA is assumed to be the first element in the array representing it, remove d_initial from the parser to reduce complexity.	2019-03-09 15:35:50 +00:00
Guido van Rossum	495da29225	bpo-35975: Support parsing earlier minor versions of Python 3 (GH-12086) This adds a `feature_version` flag to `ast.parse()` (documented) and `compile()` (hidden) that allow tweaking the parser to support older versions of the grammar. In particular if `feature_version` is 5 or 6, the hacks for the `async` and `await` keyword from PEP 492 are reinstated. (For 7 or higher, these are unconditionally treated as keywords, but they are still special tokens rather than `NAME` tokens that the parser driver recognizes.) https://bugs.python.org/issue35975	2019-03-07 12:38:08 -08:00
Serhiy Storchaka	d8b3a98c90	bpo-36187: Remove NamedStore. (GH-12167) NamedStore has been replaced with Store. The difference between NamedStore and Store is handled when precess the NamedExpr node one level upper.	2019-03-05 20:42:06 +02:00
Pablo Galindo	8bc401a55c	Clean implementation of Parser/pgen and fix some style issues (GH-12156)	2019-03-04 07:26:13 +00:00
Alex Gaynor	8589f14bbe	Remove some code which has been dead since 1994 (#12136 )	2019-03-01 23:37:34 -05:00
Pablo Galindo	1f24a719e7	bpo-35808: Retire pgen and use pgen2 to generate the parser (GH-11814) Pgen is the oldest piece of technology in the CPython repository, building it requires various #if[n]def PGEN hacks in other parts of the code and it also depends more and more on CPython internals. This commit removes the old pgen C code and replaces it for a new version implemented in pure Python. This is a modified and adapted version of lib2to3/pgen2 that can generate grammar files compatibles with the current parser. This commit also eliminates all the #ifdef and code branches related to pgen, simplifying the code and making it more maintainable. The regen-grammar step now uses $(PYTHON_FOR_REGEN) that can be any version of the interpreter, so the new pgen code maintains compatibility with older versions of the interpreter (this also allows regenerating the grammar with the current CI solution that uses Python3.5). The new pgen Python module also makes use of the Grammar/Tokens file that holds the token specification, so is always kept in sync and avoids having to maintain duplicate token definitions.	2019-03-01 15:34:44 -08:00
Pablo Galindo	b9d2e97601	Fix potential memory leak in parsetok.c (GH-11832)	2019-02-13 00:45:53 +00:00
Guido van Rossum	d2b4c19d53	bpo-35879: Fix type comment leaks (GH-11728) * Fix leak for # type: ignore * Fix the type comment leak	2019-02-01 15:28:13 -08:00
Guido van Rossum	3a32e3bf88	bpo-35766 follow-up: Kill half-support for FunctionType in PyAST_obj2mod (#11714 ) See `229874c612 (r252631862)` https://bugs.python.org/issue35766	2019-02-01 11:37:34 -08:00
Guido van Rossum	dcfcd146f8	bpo-35766: Merge typed_ast back into CPython (GH-11645)	2019-01-31 12:40:27 +01:00
Emily Morehouse	8f59ee01be	bpo-35224: PEP 572 Implementation (#10497 ) * Add tokenization of := - Add token to Include/token.h. Add token to documentation in Doc/library/token.rst. - Run `./python Lib/token.py` to regenerate Lib/token.py. - Update Parser/tokenizer.c: add case to handle `:=`. * Add initial usage of := in grammar. * Update Python.asdl to match the grammar updates. Regenerated Include/Python-ast.h and Python/Python-ast.c * Update AST and compiler files in Python/ast.c and Python/compile.c. Basic functionality, this isn't scoped properly * Regenerate Lib/symbol.py using `./python Lib/symbol.py` * Tests - Fix failing tests in test_parser.py due to changes in token numbers for internal representation * Tests - Add simple test for := token * Tests - Add simple tests for named expressions using expr and suite * Tests - Update number of levels for nested expressions to prevent stack overflow * Update symbol table to handle NamedExpr * Update Grammar to allow assignment expressions in if statements. Regenerate Python/graminit.c accordingly using `make regen-grammar` * Tests - Add additional tests for named expressions in RoundtripLegalSyntaxTestCase, based on examples and information directly from PEP 572 Note: failing tests are currently commented out (4 out of 24 tests currently fail) * Tests - Add temporary syntax test failure tests in test_parser.py Note: There is an outstanding TODO for this -- syntax tests need to be moved to a different file (presumably test_syntax.py), but this is covering what needs to be tested at the moment, and it's more convenient to run a single test for the time being * Add support for allowing assignment expressions as function argument annotations. Uncomment tests for these cases because they all pass now! * Tests - Move existing syntax tests out of test_parser.py and into test_named_expressions.py. Refactor syntax tests to use unittest * Add TargetScopeError exception to extend SyntaxError Note: This simply creates the TargetScopeError exception, it is not yet used anywhere * Tests - Update tests per PEP 572 Continue refactoring test suite: The named expression test suite now checks for any invalid cases that throw exceptions (no longer limited to SyntaxErrors), assignment tests to ensure that variables are properly assigned, and scope tests to ensure that variable availability and values are correct Note: - There are still tests that are marked to skip, as they are not yet implemented - There are approximately 300 lines of the PEP that have not yet been addressed, though these may be deferred * Documentation - Small updates to XXX/todo comments - Remove XXX from child description in ast.c - Add comment with number of previously supported nested expressions for 3.7.X in test_parser.py * Fix assert in seq_for_testlist() * Cleanup - Denote "Not implemented -- No keyword args" on failing test case. Fix PEP8 error for blank lines at beginning of test classes in test_parser.py * Tests - Wrap all file opens in `with...as` to ensure files are closed * WIP: handle f(a := 1) * Tests and Cleanup - No longer skips keyword arg test. Keyword arg test now uses a simpler test case and does not rely on an external file. Remove print statements from ast.c * Tests - Refactor last remaining test case that relied on on external file to use a simpler test case without the dependency * Tests - Add better description of remaning skipped tests. Add test checking scope when using assignment expression in a function argument * Tests - Add test for nested comprehension, testing value and scope. Fix variable name in skipped comprehension scope test * Handle restriction of LHS for named expressions - can only assign to LHS of type NAME. Specifically, restrict assignment to tuples This adds an alternative set_context specifically for named expressions, set_namedexpr_context. Thus, context is now set differently for standard assignment versus assignment for named expressions in order to handle restrictions. * Tests - Update negative test case for assigning to lambda to match new error message. Add negative test case for assigning to tuple * Tests - Reorder test cases to group invalid syntax cases and named assignment target errors * Tests - Update test case for named expression in function argument - check that result and variable are set correctly * Todo - Add todo for TargetScopeError based on Guido's comment (`2b3acd37bd (r30472562)`) * Tests - Add named expression tests for assignment operator in function arguments Note: One of two tests are skipped, as function arguments are currently treating an assignment expression inside of parenthesis as one child, which does not properly catch the named expression, nor does it count arguments properly * Add NamedStore to expr_context. Regenerate related code with `make regen-ast` * Add usage of NamedStore to ast_for_named_expr in ast.c. Update occurances of checking for Store to also handle NamedStore where appropriate * Add ste_comprehension to _symtable_entry to track if the namespace is a comprehension. Initialize ste_comprehension to 0. Set set_comprehension to 1 in symtable_handle_comprehension * s/symtable_add_def/symtable_add_def_helper. Add symtable_add_def to handle grabbing st->st_cur and passing it to symtable_add_def_helper. This now allows us to call the original code from symtable_add_def by instead calling symtable_add_def_helper with a different ste. * Refactor symtable_record_directive to take lineno and col_offset as arguments instead of stmt_ty. This allows symtable_record_directive to be used for stmt_ty and expr_ty * Handle elevating scope for named expressions in comprehensions. * Handle error for usage of named expression inside a class block * Tests - No longer skip scope tests. Add additional scope tests * Cleanup - Update error message for named expression within a comprehension within a class. Update comments. Add assert for symtable_extend_namedexpr_scope to validate that we always find at least a ModuleScope if we don't find a Class or FunctionScope * Cleanup - Add missing case for NamedStore in expr_context_name. Remove unused var in set_namedexpr_content * Refactor - Consolidate set_context and set_namedexpr_context to reduce duplicated code. Special cases for named expressions are handled by checking if ctx is NamedStore * Cleanup - Add additional use cases for ast_for_namedexpr in usage comment. Fix multiple blank lines in test_named_expressions * Tests - Remove unnecessary test case. Renumber test case function names * Remove TargetScopeError for now. Will add back if needed * Cleanup - Small comment nit for consistency * Handle positional argument check with named expression * Add TargetScopeError exception definition. Add documentation for TargetScopeError in c-api docs. Throw TargetScopeError instead of SyntaxError when using a named expression in a comprehension within a class scope * Increase stack size for parser by 200. This is a minimal change (approx. 5kb) and should not have an impact on any systems. Update parser test to allow 99 nested levels again * Add TargetScopeError to exception_hierarchy.txt for test_baseexception.py_ * Tests - Major update for named expression tests, both in test_named_expressions and test_parser - Add test for TargetScopeError - Add tests for named expressions in comprehension scope and edge cases - Add tests for named expressions in function arguments (declarations and call sites) - Reorganize tests to group them more logically * Cleanup - Remove unnecessary comment * Cleanup - Comment nitpicks * Explicitly disallow assignment expressions to a name inside parentheses, e.g.: ((x) := 0) - Add check for LHS types to detect a parenthesis then a name (see note) - Add test for this scenario - Update tests for changed error message for named assignment to a tuple (also, see note) Note: This caused issues with the previous error handling for named assignment to a LHS that contained an expression, such as a tuple. Thus, the check for the LHS of a named expression must be changed to be more specific if we wish to maintain the previous error messages * Cleanup - Wrap lines more strictly in test file * Revert "Explicitly disallow assignment expressions to a name inside parentheses, e.g.: ((x) := 0)" This reverts commit f1531400ca7d7a2d148830c8ac703f041740896d. * Add NEWS.d entry * Tests - Fix error in test_pickle.test_exceptions by adding TargetScopeError to list of exceptions * Tests - Update error message tests to reflect improved messaging convention (s/can't/cannot) * Remove cases that cannot be reached in compile.c. Small linting update. * Update Grammar/Tokens to add COLONEQUAL. Regenerate all files * Update TargetScopeError PRE_INIT and POST_INIT, as this was purposefully left out when fixing rebase conflicts * Add NamedStore back and regenerate files * Pass along line number and end col info for named expression * Simplify News entry * Fix compiler warning and explicity mark fallthrough	2019-01-24 16:49:56 -07:00
Ivan Levkivskyi	9932a22897	bpo-33416: Add end positions to Python AST (GH-11605) The majority of this PR is tediously passing `end_lineno` and `end_col_offset` everywhere. Here are non-trivial points: * It is not possible to reconstruct end positions in AST "on the fly", some information is lost after an AST node is constructed, so we need two more attributes for every AST node `end_lineno` and `end_col_offset`. * I add end position information to both CST and AST. Although it may be technically possible to avoid adding end positions to CST, the code becomes more cumbersome and less efficient. * Since the end position is not known for non-leaf CST nodes while the next token is added, this requires a bit of extra care (see `_PyNode_FinalizeEndPos`). Unless I made some mistake, the algorithm should be linear. * For statements, I "trim" the end position of suites to not include the terminal newlines and dedent (this seems to be what people would expect), for example in ```python class C: pass pass ``` the end line and end column for the class definition is (2, 8). * For `end_col_offset` I use the common Python convention for indexing, for example for `pass` the `end_col_offset` is 4 (not 3), so that `[0:4]` gives one the source code that corresponds to the node. * I added a helper function `ast.get_source_segment()`, to get source text segment corresponding to a given AST node. It is also useful for testing. An (inevitable) downside of this PR is that AST now takes almost 25% more memory. I think however it is probably justified by the benefits.	2019-01-22 11:18:22 +00:00
Anthony Sottile	995d9b9297	bpo-16806: Fix `lineno` and `col_offset` for multi-line string tokens (GH-10021)	2019-01-13 13:05:13 +09:00
Serhiy Storchaka	8ac658114d	bpo-30455: Generate all token related code and docs from Grammar/Tokens. (GH-10370) "Include/token.h", "Lib/token.py" (containing now some data moved from "Lib/tokenize.py") and new files "Parser/token.c" (containing the code moved from "Parser/tokenizer.c") and "Doc/library/token-list.inc" (included in "Doc/library/token.rst") are now generated from "Grammar/Tokens" by "Tools/scripts/generate_token.py". The script overwrites files only if needed and can be used on the read-only sources tree. "Lib/symbol.py" is now generated by "Tools/scripts/generate_symbol_py.py" instead of been executable itself. Added new make targets "regen-token" and "regen-symbol" which are now dependencies of "regen-all". The documentation contains now strings for operators and punctuation tokens.	2018-12-22 11:18:40 +02:00
Serhiy Storchaka	94cf308ee2	bpo-33306: Improve SyntaxError messages for unbalanced parentheses. (GH-6516)	2018-12-17 17:34:14 +02:00
Zackery Spytz	4c49da0cb7	bpo-35436: Add missing PyErr_NoMemory() calls and other minor bug fixes. (GH-11015) Set MemoryError when appropriate, add missing failure checks, and fix some potential leaks.	2018-12-07 12:11:30 +02:00
Victor Stinner	3bb183d7fb	bpo-35177, Python-ast.h: Fix "Yield" compiler warning (GH-10664) Partially revert commit 5f2df88b63e50d23914e97ec778861a52abdeaad: add "#undef Yield" to .c files after including Python-ast.h. Fix the warning: winbase.h(102): warning C4005: 'Yield': macro redefinition	2018-11-22 18:38:38 +01:00
Victor Stinner	621cebe81b	bpo-35081: Rename internal headers (GH-10275) Rename Include/internal/ headers: * pycore_hash.h -> pycore_pyhash.h * pycore_lifecycle.h -> pycore_pylifecycle.h * pycore_mem.h -> pycore_pymem.h * pycore_state.h -> pycore_pystate.h Add missing headers to Makefile.pre.in and PCbuild: * pycore_condvar.h. * pycore_hamt.h * pycore_pyhash.h	2018-11-12 16:53:38 +01:00
Victor Stinner	5f2df88b63	bpo-35177: Add dependencies between header files (GH-10361) * ast.h now includes Python-ast.h and node.h * parsetok.h now includes node.h and grammar.h * symtable.h now includes Python-ast.h * Modify asdl_c.py to enhance Python-ast.h: * Add #ifndef/#define Py_PYTHON_AST_H to be able to include the header twice * Add "extern { ... }" for C++ * Undefine "Yield" macro conflicting with winbase.h * Remove "#undef Yield" from C files, it's now done in Python-ast.h * Remove now useless includes in C files	2018-11-12 00:56:19 +01:00
Victor Stinner	50b48572d9	bpo-35081: Add _PyThreadState_GET() internal macro (GH-10266) If Py_BUILD_CORE is defined, the PyThreadState_GET() macro access _PyRuntime which comes from the internal pycore_state.h header. Public headers must not require internal headers. Move PyThreadState_GET() and _PyInterpreterState_GET_UNSAFE() from Include/pystate.h to Include/internal/pycore_state.h, and rename PyThreadState_GET() to _PyThreadState_GET() there. The PyThreadState_GET() macro of pystate.h is now redefined when pycore_state.h is included, to use the fast _PyThreadState_GET(). Changes: * Add _PyThreadState_GET() macro * Replace "PyThreadState_GET()->interp" with _PyInterpreterState_GET_UNSAFE() * Replace PyThreadState_GET() with _PyThreadState_GET() in internal C files (compiled with Py_BUILD_CORE defined), but keep PyThreadState_GET() in the public header files. * _testcapimodule.c: replace PyThreadState_GET() with PyThreadState_Get(); the module is not compiled with Py_BUILD_CORE defined. * pycore_state.h now requires Py_BUILD_CORE to be defined.	2018-11-01 01:51:40 +01:00
Victor Stinner	27e2d1f219	bpo-35081: Add pycore_ prefix to internal header files (GH-10263) * Rename Include/internal/ header files: * pyatomic.h -> pycore_atomic.h * ceval.h -> pycore_ceval.h * condvar.h -> pycore_condvar.h * context.h -> pycore_context.h * pygetopt.h -> pycore_getopt.h * gil.h -> pycore_gil.h * hamt.h -> pycore_hamt.h * hash.h -> pycore_hash.h * mem.h -> pycore_mem.h * pystate.h -> pycore_state.h * warnings.h -> pycore_warnings.h * PCbuild project, Makefile.pre.in, Modules/Setup: add the Include/internal/ directory to the search paths of header files. * Update includes. For example, replace #include "internal/mem.h" with #include "pycore_mem.h".	2018-11-01 00:52:28 +01:00
Serhiy Storchaka	3f22811fef	bpo-32892: Use ast.Constant instead of specific constant AST types. (GH-9445)	2018-09-27 17:42:37 +03:00
Ammar Askar	025eb98dc0	bpo-34683: Make SyntaxError column offsets consistently 1-indexed (gh-9338) Also point to start of tokens in parsing errors. Fixes bpo-34683	2018-09-24 14:12:49 -07:00
Benjamin Peterson	e502451781	closes bpo-34646: Remove PyAPI_* macros from declarations. (GH-9218)	2018-09-12 12:06:42 -07:00
Zackery Spytz	5061a74a4c	Remove unneeded PyUnicode_READY() in tokenizer.c (GH-9114)	2018-09-10 09:27:31 +03:00
Zackery Spytz	3e26e42c90	bpo-34400: Fix more undefined behavior in parsetok.c (GH-8833)	2018-08-20 20:11:40 -07:00
Zackery Spytz	7c4ab2afb1	closes bpo-34400: Fix undefined behavior in parsetok(). (GH-4439) Avoid undefined pointer arithmetic with NULL.	2018-08-14 23:27:26 -07:00
Victor Stinner	53b7d4e402	bpo-34170: Add _PyCoreConfig.bytes_warning (GH-8447) Add more fields to _PyCoreConfig: * _check_hash_pycs_mode * bytes_warning * debug * inspect * interactive * legacy_windows_fs_encoding * legacy_windows_stdio * optimization_level * quiet * unbuffered_stdio * user_site_directory * verbose * write_bytecode Changes: * Remove pymain_get_global_config() and pymain_set_global_config() which became useless. These functions have been replaced by _PyCoreConfig_GetGlobalConfig() and _PyCoreConfig_SetGlobalConfig(). * sys.flags.dont_write_bytecode value is now restricted to 1 even if -B option is specified multiple times on the command line. * PyThreadState_Clear() now uses the config from the current interpreter rather than using global Py_VerboseFlag	2018-07-25 01:37:05 +02:00
Serhiy Storchaka	aba24ff360	bpo-34084: Fix setting an error message for the "Barry as BDFL" easter egg. (GH-8262)	2018-07-23 23:41:11 +03:00
Victor Stinner	c884616390	Fix Windows compiler warning in tokenize.c (GH-8359) Fix the following warning on Windows: parser\tokenizer.c(1297): warning C4244: 'function': conversion from '__int64' to 'int', possible loss of data.	2018-07-21 03:36:06 +02:00
ValeriyaSinevich	ce75df3031	bpo-30237: Output error when ReadConsole is canceled by CancelSynchronousIo. (GH-7911)	2018-07-19 15:34:03 -07:00
Serhiy Storchaka	cf7303ed2a	bpo-33305: Improve SyntaxError for invalid numerical literals. (GH-6517)	2018-07-09 15:09:35 +03:00
Thomas A Caswell	9b9d58f0d8	bpo-31546: Fix input hook integration (GH-7978)	2018-06-28 09:29:44 -07:00
Serhiy Storchaka	a5c42284e6	bpo-33677: Fix signatures of tp_clear handlers for AST and deque. (GH-7196)	2018-05-31 07:34:34 +03:00
Serhiy Storchaka	73cbe7a01a	bpo-32911: Revert bpo-29463. (GH-7121) (GH-7197) Remove the docstring attribute of AST types and restore docstring expression as a first stmt in their body. Co-authored-by: INADA Naoki <methane@users.noreply.github.com>	2018-05-29 12:04:55 +03:00
Serhiy Storchaka	f320be77ff	bpo-32571: Avoid raising unneeded AttributeError and silencing it in C code (GH-5222) Add two new private APIs: _PyObject_LookupAttr() and _PyObject_LookupAttrId()	2018-01-25 17:49:40 +09:00
Benjamin Peterson	0a37a30037	closes bpo-32460: ensure all non-static globals have initializers (#5061 )	2017-12-31 10:04:13 -08:00
Serhiy Storchaka	598ceae876	bpo-32150: Expand tabs to spaces in C files. (#4583 )	2017-11-28 17:56:10 +02:00
Victor Stinner	9e87e7776f	bpo-32096: Remove obj and mem from _PyRuntime (#4532 ) bpo-32096, bpo-30860: Partially revert the commit 2ebc5ce42a8a9e047e790aefbf9a94811569b2b6: * Move structures back from Include/internal/mem.h to Objects/obmalloc.c * Remove _PyObject_Initialize() and _PyMem_Initialize() * Remove Include/internal/pymalloc.h * Add test_capi.test_pre_initialization_api(): Make sure that it's possible to call Py_DecodeLocale(), and then call Py_SetProgramName() with the decoded string, before Py_Initialize(). PyMem_RawMalloc() and Py_DecodeLocale() can be called again before _PyRuntimeState_Init(). Co-Authored-By: Eric Snow <ericsnowcurrently@gmail.com>	2017-11-24 12:09:24 +01:00
Victor Stinner	f2ddc6ac93	tokenizer: Remove unused tabs options (#4422 ) Remove the following fields from tok_state structure which are now used unused: * altwarning: "Issue warning if alternate tabs don't match" * alterror: "Issue error if alternate tabs don't match" * alttabsize: "Alternate tab spacing" Replace alttabsize variable with ALTTABSIZE define.	2017-11-17 01:25:47 -08:00
Victor Stinner	f7e5b56c37	bpo-32030: Split Py_Main() into subfunctions (#4399 ) * Don't use "Python runtime" anymore to parse command line options or to get environment variables: pymain_init() is now a strict separation. * Use an error message rather than "crashing" directly with Py_FatalError(). Limit the number of calls to Py_FatalError(). It prepares the code to handle errors more nicely later. * Warnings options (-W, PYTHONWARNINGS) and "XOptions" (-X) are now only added to the sys module once Python core is properly initialized. * _PyMain is now the well identified owner of some important strings like: warnings options, XOptions, and the "program name". The program name string is now properly freed at exit. pymain_free() is now responsible to free the "command" string. * Rename most methods in Modules/main.c to use a "pymain_" prefix to avoid conflits and ease debug. * Replace _Py_CommandLineDetails_INIT with memset(0) * Reorder a lot of code to fix the initialization ordering. For example, initializing standard streams now comes before parsing PYTHONWARNINGS. * Py_Main() now handles errors when adding warnings options and XOptions. * Add _PyMem_GetDefaultRawAllocator() private function. * Cleanup _PyMem_Initialize(): remove useless global constants: move them into _PyMem_Initialize(). * Call _PyRuntime_Initialize() as soon as possible: _PyRuntime_Initialize() now returns an error message on failure. * Add _PyInitError structure and following macros: * _Py_INIT_OK() * _Py_INIT_ERR(msg) * _Py_INIT_USER_ERR(msg): "user" error, don't abort() in that case * _Py_INIT_FAILED(err)	2017-11-15 15:48:08 -08:00
Serhiy Storchaka	bba2239c17	bpo-31572: Get rid of _PyObject_HasAttrId() in the ASDL parser. (#3725 ) Silence only expected AttributeError.	2017-11-11 16:41:32 +02:00
Jelle Zijlstra	ac317700ce	bpo-30406: Make async and await proper keywords (#1669 ) Per PEP 492, 'async' and 'await' should become proper keywords in 3.7.	2017-10-05 23:24:46 -04:00
Antoine Pitrou	b091bec824	bpo-31536: Avoid wholesale rebuild after `make regen-all` (#3678 ) * bpo-31536: Avoid wholesale rebuild after `make regen-all` * Add NEWS	2017-09-20 14:57:56 -07:00
Serhiy Storchaka	5d84cb368c	bpo-31464: asdl_c.py no longer emits trailing spaces in Python-ast.h. (#3568 )	2017-09-14 20:28:22 -07:00
Barry Warsaw	b2e5794870	bpo-31338 (#3374 ) * Add Py_UNREACHABLE() as an alias to abort(). * Use Py_UNREACHABLE() instead of assert(0) * Convert more unreachable code to use Py_UNREACHABLE() * Document Py_UNREACHABLE() and a few other macros.	2017-09-14 18:13:16 -07:00
Eric Snow	2ebc5ce42a	bpo-30860: Consolidate stateful runtime globals. (#3397 ) * group the (stateful) runtime globals into various topical structs * consolidate the topical structs under a single top-level _PyRuntimeState struct * add a check-c-globals.py script that helps identify runtime globals Other globals are excluded (see globals.txt and check-c-globals.py).	2017-09-07 23:51:28 -06:00
Antoine Pitrou	a6a4dc816d	bpo-31370: Remove support for threads-less builds (#3385 ) * Remove Setup.config * Always define WITH_THREAD for compatibility.	2017-09-07 18:56:24 +02:00
Eric Snow	05351c1bd8	Revert "bpo-30860: Consolidate stateful runtime globals." (#3379 ) Windows buildbots started failing due to include-related errors.	2017-09-05 21:43:08 -07:00
Eric Snow	76d5abc868	bpo-30860: Consolidate stateful runtime globals. (#2594 ) * group the (stateful) runtime globals into various topical structs * consolidate the topical structs under a single top-level _PyRuntimeState struct * add a check-c-globals.py script that helps identify runtime globals Other globals are excluded (see globals.txt and check-c-globals.py).	2017-09-05 18:26:16 -07:00
INADA Naoki	a6296d34a4	bpo-31095: fix potential crash during GC (GH-2974)	2017-08-24 14:55:17 +09:00
Yuan Chao Chou	2af565baf4	Fix a shadow-compatible-local warning (#2180 ) Change the shadowing naming, 'value' (Python-ast.c:4652), to 'val' to prevent the variables from being misused.	2017-08-04 10:53:12 -07:00
Albert-Jan Nijburg	c9ccacea3f	bpo-25324: add missing comma in Parser/tokenizer.c (GH-1910)	2017-06-01 13:51:27 -07:00
Albert-Jan Nijburg	fc354f0785	bpo-25324: copy tok_name before changing it (#1608 ) * add test to check if were modifying token * copy list so import tokenize doesnt have side effects on token * shorten line * add tokenize tokens to token.h to get them to show up in token * move ERRORTOKEN back to its previous location, and fix nitpick * copy comments from token.h automatically * fix whitespace and make more pythonic * change to fix comments from @haypo * update token.rst and Misc/NEWS * change wording * some more wording changes	2017-05-31 16:00:21 +02:00
Jim Fasarakis-Hilliard	cf1958af4c	Remove obsolete declaration in tokenizer.h (#962 )	2017-04-03 19:18:32 +03:00
Serhiy Storchaka	0b3ec19225	Use NULL rather than 0. (#778 ) There was few cases of using literal 0 instead of NULL in the context of pointers. While this was a legitimate C code, using NULL rather than 0 makes the code clearer.	2017-03-23 17:53:47 +02:00
INADA Naoki	4c78c527d2	bpo-29622: Make AST constructor to accept less than enough number of positional arguments (GH-249) bpo-29463 added optional "docstring" field to 4 AST types. While it is optional, it breaks backward compatibility because AST constructor requires number of positional argument is same to number of fields. AST types accepts empty arguments, and incomplete keyword arguments. But it's not big problem because field can be filled after creation, and checked when compiling. So stop requiring complete set of fields for positional arguments too.	2017-02-24 02:48:17 +09:00
INADA Naoki	cb41b2766d	bpo-29463: Add docstring field to some AST nodes. (#46 ) * bpo-29463: Add docstring field to some AST nodes. ClassDef, ModuleDef, FunctionDef, and AsyncFunctionDef has docstring field for now. It was first statement of there body. * fix document. thanks travis! * doc fixes	2017-02-22 16:31:59 +01:00
Berker Peksag	d2f4404bbb	Issue #28489 : Merge from 3.6	2017-02-05 04:33:11 +03:00
Berker Peksag	6f80562862	Issue #28489 : Fix comment in tokenizer.c Patch by Ryan Gonzalez.	2017-02-05 04:32:39 +03:00
INADA Naoki	fc489082c8	Issue #29369 : Use Py_IDENTIFIER in Python-ast.c	2017-01-25 22:33:43 +09:00
Victor Stinner	a5ed5f000a	Use _PyObject_CallNoArg() Replace: PyObject_CallObject(callable, NULL) with: _PyObject_CallNoArg(callable)	2016-12-06 18:45:50 +01:00
Serhiy Storchaka	06515833fe	Replaced outdated macros _PyUnicode_AsString and _PyUnicode_AsStringAndSize with PyUnicode_AsUTF8 and PyUnicode_AsUTF8AndSize.	2016-11-20 09:13:07 +02:00
Steve Dower	6c2b9d3479	Issue #28333 : Fixes off-by-one error that was adding an extra space.	2016-10-25 11:51:54 -07:00
Steve Dower	59bd34fa8a	Issue #28333 : Remove unnecessary increment.	2016-10-08 12:20:45 -07:00
Steve Dower	3cd187b9f5	Issue #28333 : Enables Unicode for ps1/ps2 and input() prompts. (Patch by Eryk Sun)	2016-10-08 12:18:16 -07:00
Serhiy Storchaka	5e80855af3	Issue #24098 : Fixed possible crash when AST is changed in process of compiling it.	2016-10-07 21:55:49 +03:00
Serhiy Storchaka	cf3806026b	Issue #24098 : Fixed possible crash when AST is changed in process of compiling it.	2016-10-07 21:51:28 +03:00
Benjamin Peterson	e2e792d98f	merge 3.5 (#28184 )	2016-09-19 22:17:16 -07:00
Benjamin Peterson	f5e8e8fc2b	merge 3.5 (#24022 )	2016-09-18 23:44:02 -07:00
Benjamin Peterson	57bda335e1	merge 3.4	2016-09-18 23:43:18 -07:00
Benjamin Peterson	26d998cfdd	properly handle the single null-byte file (closes #24022 )	2016-09-18 23:41:11 -07:00
Benjamin Peterson	9ac11a752a	properly free memory in pgen	2016-09-18 18:00:25 -07:00
Benjamin Peterson	5a715cfc57	merge 3.5 (#27981 )	2016-09-12 22:07:14 -07:00
Benjamin Peterson	35ee948fa5	restructure fp_setreadl so as to avoid refleaks (closes #27981 )	2016-09-12 22:06:58 -07:00
Brett Cannon	a721abac29	Issue #26331 : Implement the parsing part of PEP 515. Thanks to Georg Brandl for the patch.	2016-09-09 14:57:09 -07:00
Yury Selivanov	52c4e7cc84	Issue #28008 : Implement PEP 530 -- asynchronous comprehensions.	2016-09-09 10:36:01 -07:00
Yury Selivanov	f8cb8a16a3	Issue #27985 : Implement PEP 526 -- Syntax for Variable Annotations. Patch by Ivan Levkivskyi.	2016-09-08 20:50:03 -07:00
Christian Heimes	c6cc23d0b9	Skip unused value in tokenizer code In the case of an escape character, c is never read. tok_next() is used to advance the pointer. CID 1225097	2016-09-09 00:09:45 +02:00
Steve Dower	3929499914	Issue #1602 : Windows console doesn't input or print Unicode (PEP 528) Closes #17602: Adds a readline implementation for the Windows console	2016-08-30 21:22:36 -07:00
Steve Dower	940f33a50f	Issue #23524 : Finish removing _PyVerify_fd from sources	2016-09-08 11:21:54 -07:00
Benjamin Peterson	2f8bfef158	replace PY_SIZE_MAX with SIZE_MAX	2016-09-07 09:26:18 -07:00
Benjamin Peterson	ca47063998	replace Py_(u)intptr_t with the c99 standard types	2016-09-06 13:47:26 -07:00
Victor Stinner	4bb31e90f0	Fix a clang warning in grammar.c Clang is smarter than GCC and emits a warning for dead code after a function declared with __attribute__((__noreturn__)) (Py_FatalError).	2016-08-19 15:11:56 +02:00
Berker Peksag	531396c764	Issue #27336 : Fix compilation failures --without-threads	2016-06-17 13:25:01 +03:00
Serhiy Storchaka	ec39756960	Issue #22570 : Renamed Py_SETREF to Py_XSETREF.	2016-04-06 09:50:03 +03:00
Serhiy Storchaka	48842714b9	Issue #22570 : Renamed Py_SETREF to Py_XSETREF.	2016-04-06 09:45:48 +03:00
Berker Peksag	2a65ecb780	Issue #26130 : Remove redundant variable 's' from Parser/parser.c Patch by Oren Milman.	2016-03-28 00:45:28 +03:00
Benjamin Peterson	7285d520e0	remove duplicated check for fractions and complex numbers (closes #26076 ) Patch by Oren Milman.	2016-03-24 22:43:23 -07:00
Serhiy Storchaka	a051bf3afb	Issue #26581 : Use the first coding cookie on a line, not the last one.	2016-03-20 23:47:48 +02:00
Serhiy Storchaka	e431d3c9aa	Issue #26581 : Use the first coding cookie on a line, not the last one.	2016-03-20 23:36:29 +02:00
Victor Stinner	0611c26a58	On memory error, dump the memory block traceback Issue #26564: _PyObject_DebugDumpAddress() now dumps the traceback where a memory block was allocated on memory block. Use the tracemalloc module to get the traceback.	2016-03-15 22:22:13 +01:00
Victor Stinner	8a1be61849	Add more checks on the GIL Issue #10915, #15751, #26558: * PyGILState_Check() now returns 1 (success) before the creation of the GIL and after the destruction of the GIL. It allows to use the function early in Python initialization and late in Python finalization. * Add a flag to disable PyGILState_Check(). Disable PyGILState_Check() when Py_NewInterpreter() is called * Add assert(PyGILState_Check()) to: _Py_dup(), _Py_fstat(), _Py_read() and _Py_write()	2016-03-14 22:07:55 +01:00
Victor Stinner	25219f596a	Issue #26146 : remove useless code obj2ast_constant() code is baesd on obj2ast_object() which has a special case for Py_None. But in practice, we don't need to have a special case for constants. Issue noticed by Joseph Jevnik on a review.	2016-01-27 00:37:59 +01:00
Victor Stinner	f2c1aa1661	Add ast.Constant Issue #26146: Add a new kind of AST node: ast.Constant. It can be used by external AST optimizers, but the compiler does not emit directly such node. An optimizer can replace the following AST nodes with ast.Constant: * ast.NameConstant: None, False, True * ast.Num: int, float, complex * ast.Str: str * ast.Bytes: bytes * ast.Tuple if items are constants too: tuple * frozenset Update code to accept ast.Constant instead of ast.Num and/or ast.Str: * compiler * docstrings * ast.literal_eval() * Tools/parser/unparse.py	2016-01-26 00:40:57 +01:00
Serhiy Storchaka	ef1585eb9a	Issue #25923 : Added more const qualifiers to signatures of static and private functions.	2015-12-25 20:01:53 +02:00
Serhiy Storchaka	2d06e84455	Issue #25923 : Added the const qualifier to static constant arrays.	2015-12-25 19:53:18 +02:00
Serhiy Storchaka	f006940351	Issue #20440 : Massive replacing unsafe attribute setting code with special macro Py_SETREF.	2015-12-24 10:39:57 +02:00
Serhiy Storchaka	5a57ade58e	Issue #20440 : Massive replacing unsafe attribute setting code with special macro Py_SETREF.	2015-12-24 10:35:59 +02:00
Serhiy Storchaka	0304729ec4	Issue #25388 : Fixed tokenizer crash when processing undecodable source code with a null byte.	2015-11-14 15:12:04 +02:00
Serhiy Storchaka	7e2b870b85	Issue #25388 : Fixed tokenizer crash when processing undecodable source code with a null byte.	2015-11-14 15:11:17 +02:00
Serhiy Storchaka	0d441119f5	Issue #25388 : Fixed tokenizer crash when processing undecodable source code with a null byte.	2015-11-14 15:10:35 +02:00
Victor Stinner	f9827ea618	Issue #25555 : Fix parser and AST: fill lineno and col_offset of "arg" node when compiling AST from Python objects.	2015-11-06 17:01:48 +01:00
Victor Stinner	c106c68aeb	Issue #25555 : Fix parser and AST: fill lineno and col_offset of "arg" node when compiling AST from Python objects.	2015-11-06 17:01:48 +01:00
Benjamin Peterson	860c8a404a	merge 3.5 (#25502 )	2015-10-28 23:15:22 -07:00
Benjamin Peterson	669ff66c32	remove duplicated imports (closes #25502 )	2015-10-28 23:15:13 -07:00
Serhiy Storchaka	fc632e3912	Merge with 3.5.	2015-10-06 18:52:52 +03:00
Eric V. Smith	235a6f0984	Issue #24965 : Implement PEP 498 "Literal String Interpolation". Documentation is still needed, I'll open an issue for that.	2015-09-19 14:51:32 -04:00
Eric V. Smith	6408dc82fa	Fixed indentation.	2015-09-12 18:53:36 -04:00
Serhiy Storchaka	481d3af82e	Make asdl_c.py to generate Python-ast.c changed in issue #15989 .	2015-09-06 23:29:04 +03:00
Yury Selivanov	96ec934e75	Issue #24619 : Simplify async/await tokenization. This commit simplifies async/await tokenization in tokenizer.c, tokenize.py & lib2to3/tokenize.py. Previous solution was to keep a stack of async-def & def blocks, whereas the new approach is just to remember position of the outermost async-def block. This change won't bring any parsing performance improvements, but it makes the code much easier to read and validate.	2015-07-23 15:01:58 +03:00
Yury Selivanov	8fb307cd65	Issue #24619 : New approach for tokenizing async/await. This commit fixes how one-line async-defs and defs are tracked by tokenizer. It allows to correctly parse invalid code such as: >>> async def f(): ... def g(): pass ... async = 10 and valid code such as: >>> async def f(): ... async def g(): pass ... await z As a consequence, is is now possible to have one-line 'async def foo(): await ..' functions: >>> async def foo(): return await bar()	2015-07-22 13:33:45 +03:00
Yury Selivanov	8085b80c18	Issue 24226: Fix parsing of many sequential one-line 'def' statements.	2015-05-18 12:50:52 -04:00
Yury Selivanov	7544508f02	PEP 0492 -- Coroutines with async and await syntax. Issue #24017 .	2015-05-11 22:57:16 -04:00
Benjamin Peterson	025e9ebd0a	PEP 448: additional unpacking generalizations (closes #2292 ) Patch by Neil Girdhar.	2015-05-05 20:16:41 -04:00
Benjamin Peterson	273a720f87	merge 3.4 (#24022 )	2015-04-21 12:07:06 -04:00
Benjamin Peterson	d73aca769f	do not call into python api if an exception is set (#24022 )	2015-04-21 12:05:19 -04:00
Serhiy Storchaka	45ec3288d0	Removed trailing whitespaces in miscalenous files.	2015-04-03 19:42:32 +03:00
Serhiy Storchaka	a8cd4d482f	Got rid of warnings "suggest braces around empty body in an ‘else’ statement" in Parser/pgen.c.	2015-04-03 15:24:33 +03:00
Raymond Hettinger	df1b699447	Issue #22823 : Use set literals instead of creating a set from a list	2014-11-09 15:56:33 -08:00
Serhiy Storchaka	67c719b34b	Silenced some warnings about comparison between signed and unsigned integer expressions.	2014-09-05 10:10:23 +03:00
Guido van Rossum	416b516d46	Fix bootstrapping asdl -- it didn't work with Python 2.7.	2014-07-08 16:22:48 -07:00
Benjamin Peterson	3e439797ba	merge 3.4 (#21642 )	2014-06-07 12:39:51 -07:00
Benjamin Peterson	c416162302	allow the keyword else immediately after (no space) an integer (closes #21642 )	2014-06-07 12:36:39 -07:00
Eli Bendersky	5e3d338a74	Issue #19655 : Replace the ASDL parser carried with CPython The new parser does not rely on Spark (which is now removed from our repo), uses modern 3.x idioms and is significantly smaller and simpler. It generates exactly the same AST files (.h and .c), so in practice no builds should be affected.	2014-05-09 17:58:22 -07:00
Benjamin Peterson	d51374ed78	PEP 465: a dedicated infix operator for matrix multiplication (closes #21176 )	2014-04-09 23:55:56 -04:00
Martin v. Löwis	78f1e4c865	Merge with 3.3	2014-02-28 15:43:36 +01:00
Martin v. Löwis	815b41b1cd	Issue #20731 : Properly position in source code files even if they are opened in text mode. Patch by Serhiy Storchaka.	2014-02-28 15:27:29 +01:00
Benjamin Peterson	42ec031fe7	merge 3.3 (#20588 )	2014-02-10 22:41:40 -05:00
Benjamin Peterson	c2f665e721	don't put runtime values in array initializer for C89 compliance (closes #20588 )	2014-02-10 22:19:02 -05:00
Serhiy Storchaka	5940b92909	Do not reset the line number because we already set file position to correct value. (fixes error in patch for issue #18960)	2014-01-09 20:13:52 +02:00
Serhiy Storchaka	1064a13bb0	Do not reset the line number because we already set file position to correct value. (fixes error in patch for issue #18960)	2014-01-09 20:12:49 +02:00
Serhiy Storchaka	7282ff6d5b	Issue #18960 : Fix bugs with Python source code encoding in the second line. * The first line of Python script could be executed twice when the source encoding (not equal to 'utf-8') was specified on the second line. * Now the source encoding declaration on the second line isn't effective if the first line contains anything except a comment. * As a consequence, 'python -x' works now again with files with the source encoding declarations specified on the second file, and can be used again to make Python batch files on Windows. * The tokenize module now ignore the source encoding declaration on the second line if the first line contains anything except a comment. * IDLE now ignores the source encoding declaration on the second line if the first line contains anything except a comment. * 2to3 and the findnocoding.py script now ignore the source encoding declaration on the second line if the first line contains anything except a comment.	2014-01-09 18:41:59 +02:00
Serhiy Storchaka	768c16ce02	Issue #18960 : Fix bugs with Python source code encoding in the second line. * The first line of Python script could be executed twice when the source encoding (not equal to 'utf-8') was specified on the second line. * Now the source encoding declaration on the second line isn't effective if the first line contains anything except a comment. * As a consequence, 'python -x' works now again with files with the source encoding declarations specified on the second file, and can be used again to make Python batch files on Windows. * The tokenize module now ignore the source encoding declaration on the second line if the first line contains anything except a comment. * IDLE now ignores the source encoding declaration on the second line if the first line contains anything except a comment. * 2to3 and the findnocoding.py script now ignore the source encoding declaration on the second line if the first line contains anything except a comment.	2014-01-09 18:36:09 +02:00
Christian Heimes	af01f66817	Issue #16136 : Remove VMS support and VMS-related code	2013-12-21 16:19:10 +01:00
Christian Heimes	724b828e79	upcast int to size_t to silence two autological-constant-out-of-range-compare warnings with clang.	2013-12-04 08:42:46 +01:00
Victor Stinner	cad876d542	Fix a compiler warning on Windows 64-bit in parsetok.c Python parser doesn't support lines longer than INT_MAX bytes yet	2013-11-18 01:09:51 +01:00
Victor Stinner	3a8a333942	Fix compiler warnings on Windows 64-bit in grammar.c INT_MAX states and labels should be enough for everyone	2013-11-18 01:07:38 +01:00
Serhiy Storchaka	c679227e31	Issue #1772673 : The type of `char` arguments now changed to `const char`.	2013-10-19 21:03:34 +03:00
Victor Stinner	c548660af5	Issue #16742 : My fix on PyOS_StdioReadline() was incomplete, PyMem_FREE() was not patched	2013-10-19 02:40:16 +02:00
Antoine Pitrou	d01d396e7f	Issue #4555 : All exported C symbols are now prefixed with either "Py" or "_Py". ("make smelly" now clean)	2013-10-12 22:52:43 +02:00
Victor Stinner	2fe9bac4dc	Close #16742 : Fix misuse of memory allocations in PyOS_Readline() The GIL must be held to call PyMem_Malloc(), whereas PyOS_Readline() releases the GIL to read input. The result of the C callback PyOS_ReadlineFunctionPointer must now be a string allocated by PyMem_RawMalloc() or PyMem_RawRealloc() (or NULL if an error occurred), instead of a string allocated by PyMem_Malloc() or PyMem_Realloc(). Fixing this issue was required to setup a hook on PyMem_Malloc(), for example using the tracemalloc module. PyOS_Readline() copies the result of PyOS_ReadlineFunctionPointer() into a new buffer allocated by PyMem_Malloc(). So the public API of PyOS_Readline() does not change.	2013-10-10 16:18:20 +02:00
Eli Bendersky	1891cff587	Move open outside try/finally	2013-09-26 09:35:39 -07:00
Eli Bendersky	99081238e9	Don't use fancy new Python features like 'with' - some bots don't have them and can't bootstrap the parser.	2013-09-26 06:41:36 -07:00
Eli Bendersky	58fe1b1307	Normalize whitespace	2013-09-26 06:32:22 -07:00
Eli Bendersky	b788a385cd	Small fixes in Parser/asdl.py - no change in functionality. 1. Make it work when invoked directly from the command-line. It was failing due to a couple of stale function/class usages in the __main__ section. 2. Close the parsed file in the parse() function after opening it.	2013-09-26 06:31:32 -07:00
Victor Stinner	daf455554b	Issue #18571 : Implementation of the PEP 446: file descriptors and file handles are now created non-inheritable; add functions os.get/set_inheritable(), os.get/set_handle_inheritable() and socket.socket.get/set_inheritable().	2013-08-28 00:53:59 +02:00
Victor Stinner	14e461d5b9	Close #11619 : The parser and the import machinery do not encode Unicode filenames anymore on Windows.	2013-08-26 22:28:21 +02:00
Ezio Melotti	d640fe2af5	#18803 : merge with 3.3.	2013-08-26 01:33:30 +03:00
Ezio Melotti	7c4a7e6f3c	#18803 : fix more typos. Patch by Févry Thibault.	2013-08-26 01:32:56 +03:00
Antoine Pitrou	9ed5f27266	Issue #18722 : Remove uses of the "register" keyword in C code.	2013-08-13 20:18:52 +02:00
Christian Heimes	73207e03ad	Issue #18368 : PyOS_StdioReadline() no longer leaks memory when realloc() fails.	2013-08-06 16:03:33 +02:00
Christian Heimes	9ae513caa7	Issue #18368 : PyOS_StdioReadline() no longer leaks memory when realloc() fails.	2013-08-06 15:59:16 +02:00
Christian Heimes	1289565f4b	Silence warning about set but unused variable inside compile_atom() in non-debug builds	2013-07-31 23:48:04 +02:00
Christian Heimes	5e4d372524	Silence warning about set but unused variable inside compile_atom() in non-debug builds	2013-07-31 23:47:56 +02:00
Christian Heimes	b7f1b38dea	Issue #18552 : Check return value of PyArena_AddPyObject() in obj2ast_object().	2013-07-27 00:33:35 +02:00
Christian Heimes	70c94e7896	Issue #18552 : Check return value of PyArena_AddPyObject() in obj2ast_object().	2013-07-27 00:33:13 +02:00
Victor Stinner	b318990cac	(Merge 3.3) Parser/asdl_c.py: use Py_CLEAR()	2013-07-27 00:04:42 +02:00
Victor Stinner	1acc129d48	Parser/asdl_c.py: use Py_CLEAR()	2013-07-27 00:03:47 +02:00
Victor Stinner	ee4b59c0f8	(Merge 3.3) According to the PEP 7, C code must "use 4-space indents" Replace 8 spaces with 4.	2013-07-27 00:01:35 +02:00
Victor Stinner	ce72e1ce6c	According to the PEP 7, C code must "use 4-space indents" Replace 8 spaces with 4.	2013-07-27 00:00:36 +02:00
Christian Heimes	7b3902a20f	Some compilers complain about 'control reaches end of non-void function' because they don't understand that Py_FatalError() terminates the program.	2013-07-22 16:34:28 +02:00
Christian Heimes	1eb0cb12ac	Some compilers complain about 'control reaches end of non-void function' because they don't understand that Py_FatalError() terminates the program.	2013-07-22 16:34:13 +02:00
Christian Heimes	826b754e32	Add sanity check to PyGrammar_LabelRepr() in order to catch invalid tokens when debugging a new grammar. CID 715360	2013-07-22 10:30:45 +02:00
Christian Heimes	53d2dc4045	Add sanity check to PyGrammar_LabelRepr() in order to catch invalid tokens when debugging a new grammar. CID 715360	2013-07-22 10:30:14 +02:00
Victor Stinner	bdf630c4a7	Issue #18408 : Fix Python-ast.c: handle init_types() failure (ex: MemoryError)	2013-07-17 00:17:15 +02:00
Benjamin Peterson	cb2226cb69	merge 3.3	2013-07-15 20:50:25 -07:00
Benjamin Peterson	265fba40c8	move declaration to top of block	2013-07-15 20:50:22 -07:00
Benjamin Peterson	fd9c0203de	merge 3.3 (closes #18470 )	2013-07-15 20:47:47 -07:00
Benjamin Peterson	2dbfd88245	check the return value of new_string() (closes #18470 )	2013-07-15 19:15:34 -07:00
Victor Stinner	526daabf34	Issue #18408 : parsetok() must not write into stderr on memory allocation error The caller gets an error code and can raise a classic Python exception.	2013-07-11 23:17:33 +02:00
Victor Stinner	3bf5f530d9	Issue #18408 : parsetok() must not write into stderr on memory allocation error The caller gets an error code and can raise a classic Python exception.	2013-07-11 22:52:19 +02:00
Christian Heimes	22ed7fe906	Fix resource leak in parser, free node ptr CID 1028068 (#1 of 1): Resource leak (RESOURCE_LEAK) leaked_storage: Variable n going out of scope leaks the storage it points to.	2013-06-29 21:03:51 +02:00
Serhiy Storchaka	9670543a00	Issue #18038 : SyntaxError raised during compilation sources with illegal encoding now always contains an encoding name.	2013-06-09 16:53:55 +03:00
Serhiy Storchaka	3af14aaba5	Issue #18038 : SyntaxError raised during compilation sources with illegal encoding now always contains an encoding name.	2013-06-09 16:51:52 +03:00
Victor Stinner	796977360f	Issue #9566 : Fix compiler warning on Windows 64-bit	2013-06-05 00:44:00 +02:00
Benjamin Peterson	8d89c2aaba	change AST codegen to use PyModule_AddIntMacro	2013-05-20 10:28:48 -07:00
Benjamin Peterson	7654ab9ef0	placate msvc	2013-03-18 23:39:53 -07:00
Benjamin Peterson	b72406b8fa	refactor to fix refleaks	2013-03-18 23:24:41 -07:00
Benjamin Peterson	cda75be02a	unify some ast.argument's attrs; change Attribute column offset (closes #16795 ) Patch from Sven Brauch.	2013-03-18 10:48:58 -07:00
Martin v. Löwis	b26a9b10ea	Replace WaitForSingleObject with WaitForSingleObjectEx, for better WinRT compatibility.	2013-01-25 14:25:48 +01:00
Benjamin Peterson	442f20996d	create NameConstant AST class for None, True, and False literals (closes #16619 )	2012-12-06 17:41:04 -05:00
Mark Dickinson	073f067369	Issue #16546 : merge fix from 3.3	2012-11-25 14:37:43 +00:00
Mark Dickinson	ded35aeb9d	Issue #16546 : make ast.YieldFrom argument mandatory.	2012-11-25 14:36:26 +00:00
Benjamin Peterson	742b2f8d7a	make PyGrammar_LabelRepr return a const char * (closes #16369 )	2012-10-31 13:36:13 -04:00
Benjamin Peterson	d0845588b8	make _PyParser_TokenNames const	2012-10-24 08:21:52 -07:00
Matthias Klose	aee3c76acf	- Issue #16262 : fix out-of-src-tree builds, if mercurial is not installed.	2012-10-21 23:12:35 +02:00
Ezio Melotti	8a9cc526fe	#15923 : merge with 3.2.	2012-09-30 22:47:47 +03:00
Ezio Melotti	cb2916a714	#15923 : fix a mistake in asdl_c.py that resulted in a TypeError after 2801bf875a24 (see #15801 ).	2012-09-30 22:41:37 +03:00
Antoine Pitrou	ca8aa4acf6	Issue #15144 : Fix possible integer overflow when handling pointers as integer values, by using Py_uintptr_t instead of size_t. Patch by Serhiy Storchaka.	2012-09-20 20:56:47 +02:00
Georg Brandl	cc98887e45	Remove unused variables in parsetok().	2012-08-11 11:16:18 +02:00
Jesus Cea	88ca04e6a8	MERGE: Closes #15512 : Correct __sizeof__ support for parser	2012-08-03 14:29:26 +02:00
Jesus Cea	e9c5318967	Closes #15512 : Correct __sizeof__ support for parser	2012-08-03 14:28:37 +02:00
Benjamin Peterson	481ae50ccd	construct fields in the right order (closes #15517 ) Patch from Taihyun Hwang.	2012-07-31 21:41:56 -07:00
Benjamin Peterson	8107176f9b	add gc support to the AST base type (closes #15293 )	2012-07-08 11:03:46 -07:00
Antoine Pitrou	507507473e	Issue #15291 : Fix a memory leak where AST nodes where not properly deallocated.	2012-07-08 12:43:32 +02:00
Jesus Cea	035997f1a3	Issue #1677 : Unused variable warning in Non-Windows	2012-07-03 13:15:03 +02:00

... 5 6 7 8 9 ...

1170 Commits