When the lexer encounters an invalid character, it now explicitly sets the current token's type to TOKEN_ERROR and its length to 0. Previously, the token retained the previous token's type, which could cause the compiler to get stuck in an infinite loop in certain code paths. Also improves the error message for non-ASCII invalid bytes by showing the raw hex value instead of attempting to display the character, since the lexer operates on raw bytes without UTF-8 decoding. Adds regression test for issue #428 that was crashing the compiler with an out-of-bounds memory access due to this loop.
This contains the automated validation suite for the VM and built-in libraries.
-
benchmark/- Performance tests. These aren't strictly pass/fail, but let us compare performance both against other languages and against previous builds of Wren itself. -
core/- Tests for the built in core library, mainly methods on the core classes. If a bug is inwren_core.corwren_value.c, it will most likely break one of these tests. -
language/- Tests of the language itself, its grammar and runtime semantics. If a bug is inwren_compiler.corwren_vm.c, it will most likely break one of these tests. This includes tests for the syntax for the literal forms of the core classes. -
limit/- Tests for various hardcoded limits. The language doesn't officially specify these limits, but the Wren implementation has them. These tests ensure that limit behavior is well-defined and tested.