ladybird

eliott/ladybird

Fork 0

mirror of https://github.com/LadybirdBrowser/ladybird synced 2026-05-08 16:12:23 +02:00

Commit Graph

Author	SHA1	Message	Date
Andreas Kling	355fb6b825	LibWeb: Stream Rust CSS tokenizer tokens over FFI Avoid building a temporary Rust token vector before calling back into C++. The tokenizer now invokes the callback as each token is produced, while borrowing the already-filtered input for source slices. Reserve an initial C++ token capacity from the input size so the common path avoids repeated growth while appending the converted tokens. With this change, the Rust CSS tokenizer is now ~1.3x faster than the C++ CSS tokenizer at churning through all the https://vercel.com/ CSS.	2026-05-03 17:22:17 +02:00
Sam Atkins	4278194d96	LibWeb/CSS: Port the CSS Tokenizer to Rust test-css-tokenizer is updated to run both the C++ and Rust tokenizers and compare their output, to ensure they behave identically. The Parser still uses the C++ Tokenizer. The LibWeb crate, FFI layer etc are all based on the existing ones for other libraries. This is a direct AI translation to get us started, and not idiomatic Rust. Future work can be done to make it more sensible.	2026-05-03 09:49:00 +02:00

Author

SHA1

Message

Date

Andreas Kling

355fb6b825

LibWeb: Stream Rust CSS tokenizer tokens over FFI

Avoid building a temporary Rust token vector before calling back into
C++. The tokenizer now invokes the callback as each token is produced,
while borrowing the already-filtered input for source slices.

Reserve an initial C++ token capacity from the input size so the common
path avoids repeated growth while appending the converted tokens.

With this change, the Rust CSS tokenizer is now ~1.3x faster than the
C++ CSS tokenizer at churning through all the https://vercel.com/ CSS.

2026-05-03 17:22:17 +02:00

Sam Atkins

4278194d96

LibWeb/CSS: Port the CSS Tokenizer to Rust

test-css-tokenizer is updated to run both the C++ and Rust tokenizers
and compare their output, to ensure they behave identically. The Parser
still uses the C++ Tokenizer.

The LibWeb crate, FFI layer etc are all based on the existing ones for
other libraries.

This is a direct AI translation to get us started, and not idiomatic
Rust. Future work can be done to make it more sensible.

2026-05-03 09:49:00 +02:00

2 Commits