cpython

Commit Graph

Author	SHA1	Message	Date
Pablo Galindo	9f495908c5	bpo-40903: Handle multiple '=' in invalid assignment rules in the PEG parser (GH-20697) Automerge-Triggered-By: @pablogsal	2020-06-07 18:57:00 -07:00
Pablo Galindo	972ab03276	bpo-40904: Fix segfault in the new parser with f-string containing yield statements with no value (GH-20701)	2020-06-08 01:47:37 +01:00
Pablo Galindo	2e6593db00	bpo-40880: Fix invalid read in newline_in_string in pegen.c (#20666 ) * bpo-40880: Fix invalid read in newline_in_string in pegen.c * Update Parser/pegen/pegen.c Co-authored-by: Lysandros Nikolaou <lisandrosnik@gmail.com> * Add NEWS entry Co-authored-by: Lysandros Nikolaou <lisandrosnik@gmail.com>	2020-06-06 00:52:27 +01:00
Pablo Galindo	a54096e305	bpo-40883: Fix memory leak in fstring_compile_expr in parse_string.c (GH-20667)	2020-06-06 00:52:15 +01:00
Shantanu	c116c94ff1	bpo-40614: Respect feature version for f-string debug expressions (GH-20196) Co-authored-by: Lysandros Nikolaou <lisandrosnik@gmail.com> Co-authored-by: Pablo Galindo <pablogsal@gmail.com>	2020-05-27 21:30:38 +01:00
Lysandros Nikolaou	526e23f153	Refactor error handling code in Parser/pegen/pegen.c (GH-20440) Set p->error_indicator in various places, where it's needed, but it's not done. Automerge-Triggered-By: @gvanrossum	2020-05-27 09:04:11 -07:00
Pablo Galindo	404b23b85b	Fix lookahead of soft keywords in the PEG parser (GH-20436) Automerge-Triggered-By: @gvanrossum	2020-05-26 16:15:52 -07:00
Guido van Rossum	b45af1a569	Add soft keywords (GH-20370) These are like keywords but they only work in context; they are not reserved except when there is an exact match. This would enable things like match statements without reserving `match` (which would be bad for the `re.match()` function and probably lots of other places). Automerge-Triggered-By: @gvanrossum	2020-05-26 10:58:44 -07:00
Lysandros Nikolaou	f7b1e46156	bpo-38964: Print correct filename on a SyntaxError in an fstring (GH-20399) When a `SyntaxError` in the expression part of a fstring is found, the filename attribute of the `SyntaxError` is always `<fstring>`. With this commit, it gets changed to always have the name of the file the fstring resides in. Co-authored-by: Pablo Galindo <Pablogsal@gmail.com>	2020-05-26 01:32:18 +01:00
Pablo Galindo	deb4355a37	bpo-40750: Do not expand the new parser debug flags if Py_BUILD_CORE is not defined (GH-20393)	2020-05-25 20:17:12 +01:00
Pablo Galindo	800a35c623	bpo-40750: Support -d flag in the new parser (GH-20340)	2020-05-25 18:38:45 +01:00
Pablo Galindo	b23d7adfdf	Use Py_ssize_t for the column number in the PEG support code (GH-20341)	2020-05-24 06:01:34 +01:00
Lysandros Nikolaou	ae14583302	bpo-40334: Produce better error messages for non-parenthesized genexps (GH-20153) The error message, generated for a non-parenthesized generator expression in function calls, was still the generic `invalid syntax`, when the generator expression wasn't appearing as the first argument in the call. With this patch, even on input like `f(a, b, c for c in d, e)`, the correct error message gets produced.	2020-05-22 01:56:52 +01:00
Batuhan Taskaya	b8a65ec1d3	bpo-40715: Reject dict unpacking on dict comprehensions (GH-20292) Co-authored-by: Lysandros Nikolaou <lisandrosnik@gmail.com> Co-authored-by: Pablo Galindo <pablogsal@gmail.com>	2020-05-21 23:39:56 +01:00
Batuhan Taskaya	72e0aa2fd2	bpo-40176: Improve error messages for trailing comma on from import (GH-20294)	2020-05-21 21:41:58 +01:00
Pablo Galindo	ced4e5c227	Regenerate the parser (#20195 )	2020-05-18 23:47:51 +02:00
Lysandros Nikolaou	75b863aa97	bpo-40334: Reproduce error message for type comments on bare '*' in the new parser (GH-20151)	2020-05-18 20:14:47 +01:00
Lysandros Nikolaou	7b7a21bc4f	bpo-40661: Fix segfault when parsing invalid input (GH-20165) Fix segfaults when parsing very complex invalid input, like `import äˆ ð£„¯ð¢·žð±‹á”€ð””ð‘©±å®ä±¬ð©¾\nð—¶½`. Co-authored-by: Guido van Rossum <guido@python.org> Co-authored-by: Pablo Galindo <pablogsal@gmail.com>	2020-05-18 18:32:03 +01:00
Lysandros Nikolaou	2c8cd06afe	bpo-40334: Improvements to error-handling code in the PEG parser (GH-20003) The following improvements are implemented in this commit: - `p->error_indicator` is set, in case malloc or realloc fail. - Avoid memory leaks in the case that realloc fails. - Call `PyErr_NoMemory()` instead of `PyErr_Format()`, because it requires no memory. Co-authored-by: Pablo Galindo <Pablogsal@gmail.com>	2020-05-17 04:19:23 +01:00
Pablo Galindo	16ab07063c	bpo-40334: Correctly identify invalid target in assignment errors (GH-20076) Co-authored-by: Lysandros Nikolaou <lisandrosnik@gmail.com>	2020-05-15 02:04:52 +01:00
Lysandros Nikolaou	ce21cfca7b	bpo-40618: Disallow invalid targets in augassign and except clauses (GH-20083) This commit fixes the new parser to disallow invalid targets in the following scenarios: - Augmented assignments must only accept a single target (Name, Attribute or Subscript), but no tuples or lists. - `except` clauses should only accept a single `Name` as a target. Co-authored-by: Pablo Galindo <Pablogsal@gmail.com>	2020-05-14 21:13:50 +01:00
Pablo Galindo	bcc3036095	bpo-40619: Correctly handle error lines in programs without file mode (GH-20090)	2020-05-14 21:11:48 +01:00
Lysandros Nikolaou	a15c9b3a05	bpo-40334: Always show the caret on SyntaxErrors (GH-20050) This commit fixes SyntaxError locations when the caret is not displayed, by doing the following: - `col_number` always gets set to the location of the offending node/expr. When no caret is to be displayed, this gets achieved by setting the object holding the error line to None. - Introduce a new function `_PyPegen_raise_error_known_location`, which can be called, when an arbitrary `lineno`/`col_offset` needs to be passed. This function then gets used in the grammar (through some new macros and inline functions) so that SyntaxError locations of the new parser match that of the old.	2020-05-13 20:36:27 +01:00
Serhiy Storchaka	74ea6b5a75	bpo-40593: Improve syntax errors for invalid characters in source code. (GH-20033)	2020-05-12 12:42:04 +03:00
Shantanu	27c0d9b54a	bpo-40334: produce specialized errors for invalid del targets (GH-19911)	2020-05-11 14:53:58 -07:00
Pablo Galindo	5b956ca42d	bpo-40585: Normalize errors messages in codeop when comparing them (GH-20030) With the new parser, the error message contains always the trailing newlines, causing the comparison of the repr of the error messages in codeop to fail. This commit makes the new parser mirror the old parser's behaviour regarding trailing newlines.	2020-05-11 01:41:26 +01:00
Pablo Galindo	ac7a92cc0a	bpo-40334: Avoid collisions between parser variables and grammar variables (GH-19987) This is for the C generator: - Disallow rule and variable names starting with `_` - Rename most local variable names generated by the parser to start with `_` Exceptions: - Renaming `p` to `_p` will be a separate PR - There are still some names that might clash, e.g. - anything starting with `Py` - C reserved words (`if` etc.) - Macros like `EXTRA` and `CHECK`	2020-05-09 21:34:50 -07:00
Pablo Galindo	db9163ceef	bpo-40555: Check for p->error_indicator in loop rules after the main loop is done (GH-19986)	2020-05-08 03:38:44 +01:00
Lysandros Nikolaou	4638c64295	bpo-40334: Error message for invalid default args in function call (GH-19973) When parsing something like `f(g()=2)`, where the name of a default arg is not a NAME, but an arbitrary expression, a specialised error message is emitted.	2020-05-07 11:44:06 +01:00
Lysandros Nikolaou	2f37c355ab	bpo-40334: Fix error location upon parsing an invalid string literal (GH-19962) When parsing a string with an invalid escape, the old parser used to point to the beginning of the invalid string. This commit changes the new parser to match that behaviour, since it's currently pointing to the end of the string (or to be more precise, to the beginning of the next token).	2020-05-07 11:37:51 +01:00
Pablo Galindo	470aac4d8e	bpo-40334: Generate comments in the parser code to improve debugging (GH-19966)	2020-05-06 23:14:43 +01:00
Pablo Galindo	99db2a1db7	bpo-40334: Allow trailing comma in parenthesised context managers (GH-19964)	2020-05-06 22:54:34 +01:00
Lysandros Nikolaou	999ec9ab6a	bpo-40334: Add type to the assignment rule in the grammar file (GH-19963)	2020-05-06 19:11:04 +01:00
Lysandros Nikolaou	846d8b28ab	bpo-40246: Revert reporting of invalid string prefixes (GH-19888) Due to backwards compatibility concerns regarding keywords immediately followed by a string without whitespace between them (like in `bg="#d00" if clear else"#fca"`) will fail to parse, commit `41d5b94af4` has to be reverted.	2020-05-04 12:32:18 +01:00
Lysandros Nikolaou	e10e7c771b	bpo-40334: Spacialized error message for invalid args after bare '' (GH-19865) When parsing things like `def f(): pass` the old parser used to output `SyntaxError: named arguments must follow bare *`, which the new parser wasn't able to do.	2020-05-04 11:58:31 +01:00
Shantanu	c3f001461d	bpo-40491: Fix typo in syntax error for numeric literals (GH-19893)	2020-05-04 11:13:30 +03:00
Shantanu	603d354626	bpo-40493: fix function type comment parsing (GH-19894) The grammar for func_type_input rejected things like `(*t1) ->t2`. This fixes that. Automerge-Triggered-By: @gvanrossum	2020-05-03 22:08:14 -07:00
Lysandros Nikolaou	7f06af684a	bpo-40334: Set error_indicator in _PyPegen_raise_error (GH-19887) Due to PyErr_Occurred not being called at the beginning of each rule, we need to set the error indicator, so that rules do not get expanded after an exception has been thrown	2020-05-04 01:20:09 +01:00
Lysandros Nikolaou	03b7642265	bpo-40334: Make the PyPegen* and PyParser* APIs more consistent (GH-19839) This commit makes both APIs more consistent by doing the following: - Remove the `PyPegen_CodeObjectFrom` functions, which weren't used and will probably not be needed. Functions like `Py_CompileStringObject` can be used instead. - Include a `const char filename` parameter in `PyPegen_ASTFromString`. - Rename `PyPegen_ASTFromFile` to `PyPegen_ASTFromFilename`, because its signature is not the same with `PyParser_ASTFromFile`.	2020-05-01 18:30:51 +01:00
Guido van Rossum	d9d6eadf00	Ensure that tok->type_comments is set on every path (GH-19828)	2020-05-01 17:42:32 +01:00
Guido van Rossum	3941d9700b	bpo-40334: Refactor lambda_parameters similar to parameters (GH-19830)	2020-05-01 17:42:03 +01:00
Pablo Galindo	d955241469	bpo-40334: Correct return value of func_type_comment (GH-19833)	2020-05-01 08:32:09 -07:00
Batuhan Taskaya	76c1b4d5c5	bpo-40334: Improve column offsets for thrown syntax errors by Pegen (GH-19782)	2020-05-01 14:13:43 +01:00
Pablo Galindo	b796b3fb48	bpo-40334: Simplify type handling in the PEG c_generator (GH-19818)	2020-05-01 12:32:26 +01:00
Lysandros Nikolaou	3e0a6f37df	bpo-40334: Add support for feature_version in new PEG parser (GH-19827) `ast.parse` and `compile` support a `feature_version` parameter that tells the parser to parse the input string, as if it were written in an older Python version. The `feature_version` is propagated to the tokenizer, which uses it to handle the three different stages of support for `async` and `await`. Additionally, it disallows the following at parser level: - The '@' operator in < 3.5 - Async functions in < 3.5 - Async comprehensions in < 3.6 - Underscores in numeric literals in < 3.6 - Await expression in < 3.5 - Variable annotations in < 3.6 - Async for-loops in < 3.5 - Async with-statements in < 3.5 - F-strings in < 3.6 Closes we-like-parsers/cpython#124.	2020-04-30 20:27:52 -07:00
Guido van Rossum	c001c09e90	bpo-40334: Support type comments (GH-19780) This implements full support for # type: <type> comments, # type: ignore <stuff> comments, and the func_type parsing mode for ast.parse() and compile(). Closes https://github.com/we-like-parsers/cpython/issues/95. (For now, you need to use the master branch of mypy, since another issue unique to 3.9 had to be fixed there, and there's no mypy release yet.) The only thing missing is `feature_version=N`, which is being tracked in https://github.com/we-like-parsers/cpython/issues/124.	2020-04-30 12:12:19 -07:00
Pablo Galindo	4db245ee9d	bpo-40334: refactor and cleanup for the PEG generators (GH-19775)	2020-04-29 10:42:21 +01:00
Lysandros Nikolaou	6d65087655	bpo-40334: Disallow invalid single statements in the new parser (GH-19774) After parsing is done in single statement mode, the tokenizer buffer has to be checked for additional lines and a `SyntaxError` must be raised, in case there are any. Co-authored-by: Pablo Galindo <Pablogsal@gmail.com>	2020-04-29 02:42:27 +01:00
Pablo Galindo	2208134918	bpo-40334: Explicitly cast to int in pegen.c to fix a compiler warning (GH-19779)	2020-04-29 02:04:06 +01:00
Lysandros Nikolaou	37af21b667	bpo-40334: Fix shifting of nested f-strings in the new parser (GH-19771) `JoinedStr`s and `FormattedValue also needs to be shifted, in order to correctly compute the location information of nested f-strings.	2020-04-29 01:43:50 +01:00

1 2

60 Commits