Python/ruff - ruff - Gitea: Git with a cup of tea

mirror of https://github.com/astral-sh/ruff synced 2026-01-20 21:10:48 -05:00

Author	SHA1	Message	Date
Dhruv Manilawala	4f49e918a9	Bump version to v0.4.9 (#11872 )	2024-06-14 20:36:22 +05:30
Micha Reiser	d4dd96d1f4	red-knot: `source_text`, `line_index`, and `parsed_module` queries (#11822 )	2024-06-13 07:37:02 +00:00
Micha Reiser	efbf7b14b5	red-knot[salsa part 2]: Setup semantic DB and Jar (#11837 ) Co-authored-by: Alex Waygood <Alex.Waygood@Gmail.com>	2024-06-13 08:00:51 +01:00
Alex Waygood	4ed3aed8d3	[red-knot] Add a parser for typeshed's VERSIONS file (#11836 )	2024-06-12 11:44:45 +00:00
Micha Reiser	93973b96cb	red-knot: `VfsFile` input ingredient and a `Vfs` (#11802 )	2024-06-12 07:06:15 +00:00
Jane Lewis	507f5c1137	`ruff server`: Tracing system now respects log level and trace level, with options to log to a file (#11747 ) ## Summary Fixes #10968. Fixes #11545. The server's tracing system has been rewritten from the ground up. The server now has trace level and log level settings which restrict the tracing events and spans that get logged. * A `logLevel` setting has been added, which lets a user set the log level. By default, it is set to `"info"`. * A `logFile` setting has also been added, which lets the user supply an optional file to send tracing output (it does not have to exist as a file yet). By default, if this is unset, tracing output will be sent to `stderr`. * A `$/setTrace` handler has also been added, and we also set the trace level from the initialization options. For editors without direct support for tracing, the environment variable `RUFF_TRACE` can override the trace level. * Small changes have been made to how we display tracing output. We no longer use `tracing-tree`, and instead use `tracing_subscriber::fmt::Layer` to format output. Thread names are now included in traces, and I've made some adjustment to thread worker names to be more useful. ## Test Plan In VS Code, with `ruff.trace.server` set to its default value, no logs from Ruff should appear. After changing `ruff.trace.server` to either `messages` or `verbose`, you should see log messages at `info` level or higher appear in Ruff's output: <img width="1005" alt="Screenshot 2024-06-10 at 10 35 04 AM" src="https://github.com/astral-sh/ruff/assets/19577865/6050d107-9815-4bd2-96d0-e86f096a57f5"> In Helix, by default, no logs from Ruff should appear. To set the trace level in Helix, you'll need to modify your language configuration as follows: ```toml [language-server.ruff] command = "/Users/jane/astral/ruff/target/debug/ruff" args = ["server", "--preview"] environment = { "RUFF_TRACE" = "messages" } ``` After doing this, logs of `info` level or higher should be visible in Helix: <img width="1216" alt="Screenshot 2024-06-10 at 10 39 26 AM" src="https://github.com/astral-sh/ruff/assets/19577865/8ff88692-d3f7-4fd1-941e-86fb338fcdcc"> You can use `:log-open` to quickly open the Helix log file. In Neovim, by default, no logs from Ruff should appear. To set the trace level in Neovim, you'll need to modify your configuration as follows: ```lua require('lspconfig').ruff.setup { cmd = {"/path/to/debug/executable", "server", "--preview"}, cmd_env = { RUFF_TRACE = "messages" } } ``` You should see logs appear in `:LspLog` that look like the following: <img width="1490" alt="Screenshot 2024-06-11 at 11 24 01 AM" src="https://github.com/astral-sh/ruff/assets/19577865/576cd5fa-03cf-477a-b879-b29a9a1200ff"> You can adjust `logLevel` and `logFile` in `settings`: ```lua require('lspconfig').ruff.setup { cmd = {"/path/to/debug/executable", "server", "--preview"}, cmd_env = { RUFF_TRACE = "messages" }, settings = { logLevel = "debug", logFile = "your/log/file/path/log.txt" } } ``` The `logLevel` and `logFile` can also be set in Helix like so: ```toml [language-server.ruff.config.settings] logLevel = "debug" logFile = "your/log/file/path/log.txt" ``` Even if this log file does not exist, it should now be created and written to after running the server: <img width="1148" alt="Screenshot 2024-06-10 at 10 43 44 AM" src="https://github.com/astral-sh/ruff/assets/19577865/ab533cf7-d5ac-4178-97f1-e56da17450dd">	2024-06-11 11:29:47 -07:00
renovate[bot]	e78b9dc7fe	Update Rust crate strum_macros to v0.26.4 (#11814 ) \	2024-06-09 21:47:56 -04:00
renovate[bot]	8b9bbc0c84	Update Rust crate toml to v0.8.14 (#11815 )	2024-06-10 01:47:45 +00:00
renovate[bot]	87ea06e360	Update Rust crate regex to v1.10.5 (#11813 )	2024-06-10 01:47:23 +00:00
renovate[bot]	74246f4acc	Update Rust crate clap to v4.5.6 (#11812 )	2024-06-10 01:47:06 +00:00
Dhruv Manilawala	549cc1e437	Build `CommentRanges` outside the parser (#11792 ) ## Summary This PR updates the parser to remove building the `CommentRanges` and instead it'll be built by the linter and the formatter when it's required. For the linter, it'll be built and owned by the `Indexer` while for the formatter it'll be built from the `Tokens` struct and passed as an argument. ## Test Plan `cargo insta test`	2024-06-09 09:55:17 +00:00
Alex Waygood	37d8de3316	[red-knot] Include vendored typeshed stubs as a zipfile in the Ruff binary (#11779 ) Co-authored-by: Micha Reiser <micha@reiser.io> Co-authored-by: Carl Meyer <carl@astral.sh>	2024-06-07 15:00:36 +00:00
Dhruv Manilawala	d22f3402e1	Remove `result_like` dependency (#11793 ) ## Summary This PR removes the `result-like` dependency and instead implement the required functionality. The motivation being that `noqa.is_enabled()` is easier to read than `noqa.into()`. For context, I was just trying to understand the syntax error workflow and I saw these flags which were being converted via `into`. I always find `into` confusing because you never know what's it being converted into unless you know the type. Later realized that it's just a boolean flag. After removing the usages from these two flags, it turns out that the dependency is only being used in one rule so I thought to remove that as well. ## Test Plan `cargo insta test`	2024-06-07 11:53:22 +05:30
Alex Waygood	303ef02f93	[red-knot] Encapsulate module resolution logic in `module.rs` (#11767 )	2024-06-06 14:31:09 +00:00
Micha Reiser	5806bc915d	Fix formatter instability for lines only consisting of zero-width characters (#11748 )	2024-06-05 17:55:14 +02:00
Dhruv Manilawala	a8cf7096ff	Bump version to v0.4.8 (#11755 ) Co-authored-by: Alex Waygood <Alex.Waygood@Gmail.com>	2024-06-05 20:51:31 +05:30
Micha Reiser	64165bee43	red-knot: Use `parse_unchecked` to get all parse errors (#11725 )	2024-06-04 06:04:48 +00:00
Dhruv Manilawala	a58bde6958	Remove less used parser dependencies (#11718 ) ## Summary This PR removes the following dependencies from the `ruff_python_parser` crate: * `anyhow` (moved to dev dependencies) * `is-macro` * `itertools` The main motivation is that they aren't used much. Additionally, it updates the return type of `parse_type_annotation` to use a more specific `ParseError` instead of the generic `anyhow::Error`. ## Test Plan `cargo insta test`	2024-06-03 13:08:24 +00:00
Dhruv Manilawala	bf5b62edac	Maintain synchronicity between the lexer and the parser (#11457 ) ## Summary This PR updates the entire parser stack in multiple ways: ### Make the lexer lazy * https://github.com/astral-sh/ruff/pull/11244 * https://github.com/astral-sh/ruff/pull/11473 Previously, Ruff's lexer would act as an iterator. The parser would collect all the tokens in a vector first and then process the tokens to create the syntax tree. The first task in this project is to update the entire parsing flow to make the lexer lazy. This includes the `Lexer`, `TokenSource`, and `Parser`. For context, the `TokenSource` is a wrapper around the `Lexer` to filter out the trivia tokens[^1]. Now, the parser will ask the token source to get the next token and only then the lexer will continue and emit the token. This means that the lexer needs to be aware of the "current" token. When the `next_token` is called, the current token will be updated with the newly lexed token. The main motivation to make the lexer lazy is to allow re-lexing a token in a different context. This is going to be really useful to make the parser error resilience. For example, currently the emitted tokens remains the same even if the parser can recover from an unclosed parenthesis. This is important because the lexer emits a `NonLogicalNewline` in parenthesized context while a normal `Newline` in non-parenthesized context. This different kinds of newline is also used to emit the indentation tokens which is important for the parser as it's used to determine the start and end of a block. Additionally, this allows us to implement the following functionalities: 1. Checkpoint - rewind infrastructure: The idea here is to create a checkpoint and continue lexing. At a later point, this checkpoint can be used to rewind the lexer back to the provided checkpoint. 2. Remove the `SoftKeywordTransformer` and instead use lookahead or speculative parsing to determine whether a soft keyword is a keyword or an identifier 3. Remove the `Tok` enum. The `Tok` enum represents the tokens emitted by the lexer but it contains owned data which makes it expensive to clone. The new `TokenKind` enum just represents the type of token which is very cheap. This brings up a question as to how will the parser get the owned value which was stored on `Tok`. This will be solved by introducing a new `TokenValue` enum which only contains a subset of token kinds which has the owned value. This is stored on the lexer and is requested by the parser when it wants to process the data. For example: `8196720f80/crates/ruff_python_parser/src/parser/expression.rs (L1260-L1262)` [^1]: Trivia tokens are `NonLogicalNewline` and `Comment` ### Remove `SoftKeywordTransformer` * https://github.com/astral-sh/ruff/pull/11441 * https://github.com/astral-sh/ruff/pull/11459 * https://github.com/astral-sh/ruff/pull/11442 * https://github.com/astral-sh/ruff/pull/11443 * https://github.com/astral-sh/ruff/pull/11474 For context, https://github.com/RustPython/RustPython/pull/4519/files#diff-5de40045e78e794aa5ab0b8aacf531aa477daf826d31ca129467703855408220 added support for soft keywords in the parser which uses infinite lookahead to classify a soft keyword as a keyword or an identifier. This is a brilliant idea as it basically wraps the existing Lexer and works on top of it which means that the logic for lexing and re-lexing a soft keyword remains separate. The change here is to remove `SoftKeywordTransformer` and let the parser determine this based on context, lookahead and speculative parsing. * Context: The transformer needs to know the position of the lexer between it being at a statement position or a simple statement position. This is because a `match` token starts a compound statement while a `type` token starts a simple statement. The parser already knows this. * Lookahead: Now that the parser knows the context it can perform lookahead of up to two tokens to classify the soft keyword. The logic for this is mentioned in the PR implementing it for `type` and `match soft keyword. * Speculative parsing: This is where the checkpoint - rewind infrastructure helps. For `match` soft keyword, there are certain cases for which we can't classify based on lookahead. The idea here is to create a checkpoint and keep parsing. Based on whether the parsing was successful and what tokens are ahead we can classify the remaining cases. Refer to #11443 for more details. If the soft keyword is being parsed in an identifier context, it'll be converted to an identifier and the emitted token will be updated as well. Refer `8196720f80/crates/ruff_python_parser/src/parser/expression.rs (L487-L491)`. The `case` soft keyword doesn't require any special handling because it'll be a keyword only in the context of a match statement. ### Update the parser API * https://github.com/astral-sh/ruff/pull/11494 * https://github.com/astral-sh/ruff/pull/11505 Now that the lexer is in sync with the parser, and the parser helps to determine whether a soft keyword is a keyword or an identifier, the lexer cannot be used on its own. The reason being that it's not sensitive to the context (which is correct). This means that the parser API needs to be updated to not allow any access to the lexer. Previously, there were multiple ways to parse the source code: 1. Passing the source code itself 2. Or, passing the tokens Now that the lexer and parser are working together, the API corresponding to (2) cannot exists. The final API is mentioned in this PR description: https://github.com/astral-sh/ruff/pull/11494. ### Refactor the downstream tools (linter and formatter) * https://github.com/astral-sh/ruff/pull/11511 * https://github.com/astral-sh/ruff/pull/11515 * https://github.com/astral-sh/ruff/pull/11529 * https://github.com/astral-sh/ruff/pull/11562 * https://github.com/astral-sh/ruff/pull/11592 And, the final set of changes involves updating all references of the lexer and `Tok` enum. This was done in two-parts: 1. Update all the references in a way that doesn't require any changes from this PR i.e., it can be done independently * https://github.com/astral-sh/ruff/pull/11402 * https://github.com/astral-sh/ruff/pull/11406 * https://github.com/astral-sh/ruff/pull/11418 * https://github.com/astral-sh/ruff/pull/11419 * https://github.com/astral-sh/ruff/pull/11420 * https://github.com/astral-sh/ruff/pull/11424 2. Update all the remaining references to use the changes made in this PR For (2), there were various strategies used: 1. Introduce a new `Tokens` struct which wraps the token vector and add methods to query a certain subset of tokens. These includes: 1. `up_to_first_unknown` which replaces the `tokenize` function 2. `in_range` and `after` which replaces the `lex_starts_at` function where the former returns the tokens within the given range while the latter returns all the tokens after the given offset 2. Introduce a new `TokenFlags` which is a set of flags to query certain information from a token. Currently, this information is only limited to any string type token but can be expanded to include other information in the future as needed. https://github.com/astral-sh/ruff/pull/11578 3. Move the `CommentRanges` to the parsed output because this information is common to both the linter and the formatter. This removes the need for `tokens_and_ranges` function. ## Test Plan - [x] Update and verify the test snapshots - [x] Make sure the entire test suite is passing - [x] Make sure there are no changes in the ecosystem checks - [x] Run the fuzzer on the parser - [x] Run this change on dozens of open-source projects ### Running this change on dozens of open-source projects Refer to the PR description to get the list of open source projects used for testing. Now, the following tests were done between `main` and this branch: 1. Compare the output of `--select=E999` (syntax errors) 2. Compare the output of default rule selection 3. Compare the output of `--select=ALL` Conclusion: all output were same ## What's next? The next step is to introduce re-lexing logic and update the parser to feed the recovery information to the lexer so that it can emit the correct token. This moves us one step closer to having error resilience in the parser and provides Ruff the possibility to lint even if the source code contains syntax errors.	2024-06-03 18:23:50 +05:30
renovate[bot]	ded010cf9c	Update Rust crate tracing-tree to v0.3.1 (#11703 )	2024-06-02 21:51:13 -04:00
renovate[bot]	436dc18b15	Update Rust crate libcst to v1.4.0 (#11707 )	2024-06-03 01:05:32 +00:00
renovate[bot]	9599bd7622	Update Rust crate itertools to 0.13.0 (#11706 )	2024-06-03 01:05:17 +00:00
renovate[bot]	ec3f523924	Update Rust crate insta to v1.39.0 (#11705 )	2024-06-03 01:04:26 +00:00
renovate[bot]	010434015e	Update Rust crate proc-macro2 to v1.0.85 (#11700 )	2024-06-03 01:03:31 +00:00
renovate[bot]	25131da2c3	Update Rust crate toml to v0.8.13 (#11702 )	2024-06-02 21:03:09 -04:00
renovate[bot]	712783825d	Update Rust crate strum_macros to v0.26.3 (#11701 )	2024-06-02 21:03:03 -04:00
Charlie Marsh	1ad5f9c038	Bump version to v0.4.7 (#11646 )	2024-05-31 16:30:36 -04:00
Charlie Marsh	49a5a9ccc2	Bump version to v0.4.6 (#11585 )	2024-05-28 15:10:53 -04:00
Charlie Marsh	16acd4913f	Remove some unused `pub` functions (#11576 ) ## Summary I left anything in `red-knot`, any `with_` methods, etc.	2024-05-28 09:56:51 -04:00
Charlie Marsh	34a5063aa2	Respect excludes in `ruff server` configuration discovery (#11551 ) ## Summary Right now, we're discovering configuration files even within (e.g.) virtual environments, because we're recursing without respecting the `exclude` field on parent configuration. Closes https://github.com/astral-sh/ruff-vscode/issues/478. ## Test Plan Installed Pandas; verified that I saw no warnings: ![Screenshot 2024-05-26 at 8 09 05 PM](https://github.com/astral-sh/ruff/assets/1309177/dcf4115c-d7b3-453b-b7c7-afdd4804d6f5)	2024-05-27 16:59:46 +00:00
renovate[bot]	5dcde88099	Update Rust crate thiserror to v1.0.61 (#11561 )	2024-05-27 00:33:54 +00:00
renovate[bot]	40bfae4f99	Update Rust crate syn to v2.0.66 (#11560 )	2024-05-27 00:21:44 +00:00
renovate[bot]	7b064b25b2	Update Rust crate mimalloc to v0.1.42 (#11554 )	2024-05-26 20:21:39 -04:00
renovate[bot]	9993115f63	Update Rust crate smol_str to v0.2.2 (#11559 )	2024-05-26 20:21:25 -04:00
renovate[bot]	f0a21c9161	Update Rust crate serde to v1.0.203 (#11558 )	2024-05-26 20:21:19 -04:00
renovate[bot]	f26c155de5	Update Rust crate schemars to v0.8.21 (#11557 )	2024-05-26 20:21:13 -04:00
renovate[bot]	c3fa826b0a	Update Rust crate parking_lot to v0.12.3 (#11555 )	2024-05-26 20:21:03 -04:00
renovate[bot]	8b69794f1d	Update Rust crate libc to v0.2.155 (#11553 )	2024-05-26 20:20:47 -04:00
renovate[bot]	4e7c84df1d	Update Rust crate anyhow to v1.0.86 (#11552 )	2024-05-26 20:20:38 -04:00
Jane Lewis	550aa871d3	Bump version to `v0.4.5` (#11502 )	2024-05-23 01:09:01 +00:00
Jane Lewis	b0731ef9cb	`ruff server`: Support Jupyter Notebook (`.ipynb`) files (#11206 ) ## Summary Closes https://github.com/astral-sh/ruff/issues/10858. `ruff server` now supports `.ipynb` (aka Jupyter Notebook) files. Extensive internal changes have been made to facilitate this, which I've done some work to contextualize with documentation and an pre-review that highlights notable sections of the code. `.ipynb` cells should behave similarly to `.py` documents, with one major exception. The format command `ruff.applyFormat` will only apply to the currently selected notebook cell - if you want to format an entire notebook document, use `Format Notebook` from the VS Code context menu. ## Test Plan The VS Code extension does not yet have Jupyter Notebook support enabled, so you'll first need to enable it manually. To do this, checkout the `pre-release` branch and modify `src/common/server.ts` as follows: Before: ![Screenshot 2024-05-13 at 10 59 06 PM](https://github.com/astral-sh/ruff/assets/19577865/c6a3c604-c405-4968-b8a2-5d670de89172) After: ![Screenshot 2024-05-13 at 10 58 24 PM](https://github.com/astral-sh/ruff/assets/19577865/94ab2e3d-0609-448d-9c8c-cd07c69a513b) I recommend testing this PR with large, complicated notebook files. I used notebook files from [this popular repository](https://github.com/jakevdp/PythonDataScienceHandbook/tree/master/notebooks) in my preliminary testing. The main thing to test is ensuring that notebook cells behave the same as Python documents, besides the aforementioned issue with `ruff.applyFormat`. You should also test adding and deleting cells (in particular, deleting all the code cells and ensure that doesn't break anything), changing the kind of a cell (i.e. from markup -> code or vice versa), and creating a new notebook file from scratch. Finally, you should also test that source actions work as expected (and across the entire notebook). Note: `ruff.applyAutofix` and `ruff.applyOrganizeImports` are currently broken for notebook files, and I suspect it has something to do with https://github.com/astral-sh/ruff/issues/11248. Once this is fixed, I will update the test plan accordingly. --------- Co-authored-by: nolan <nolan.king90@gmail.com>	2024-05-21 22:29:30 +00:00
Charlie Marsh	6cec82fff8	Get `cargo shear` passing (#11392 ) ## Summary Remove some unused dependencies, add a few ignores.	2024-05-13 01:56:24 +00:00
renovate[bot]	7c824faa88	Update Rust crate thiserror to v1.0.60 (#11390 )	2024-05-13 00:36:08 +00:00
renovate[bot]	12da5968a0	Update Rust crate serde_json to v1.0.117 (#11388 )	2024-05-13 00:35:46 +00:00
renovate[bot]	a747b3f2a1	Update Rust crate syn to v2.0.63 (#11389 )	2024-05-13 00:35:23 +00:00
renovate[bot]	01a0e6cc7e	Update Rust crate serde to v1.0.201 (#11387 )	2024-05-13 00:34:34 +00:00
renovate[bot]	a8b06537c7	Update Rust crate anyhow to v1.0.83 (#11384 )	2024-05-13 00:34:00 +00:00
renovate[bot]	7b8fe25d32	Update Rust crate schemars to v0.8.19 (#11386 )	2024-05-13 00:33:29 +00:00
renovate[bot]	a50416a6d7	Update Rust crate proc-macro2 to v1.0.82 (#11385 )	2024-05-13 00:33:05 +00:00
Alex Waygood	3e8878a1c8	Bump version to v0.4.4 (#11352 )	2024-05-09 17:00:46 +00:00

1 2 3 4 5 ...

954 Commits