A RewriteInRustBench would be unironically useful at this point since all the main agents can write it reasonably well despite its relative scarcity in the input data.
From DARPA’s “Translating all C to Rust” TRACTOR program:
https://www.ll.mit.edu/r-d/projects/translating-all-c-rust-t...
I wouldn't be surprised if we're already at the point of more LLM-written Rust than hand-written. Models training off models
Have each agent rewrite openssl in $lang and call it the RollYourOwnCrypto bench.
All the main agents can write Jai code reasonably well despite being even more scarce in input data!
Rust is the best language for LLMs b/c it gives by far the best debug messages. Just tons of verifiable reward signal for post-training. Even the most rudimentary LLMs can school me on idiomatic Rust