benchmarks.wiki / Public workspace

AssistantBench / f94fa0a81aedde2ba39236a7f64988dbfa92a41a19f26f0caca81f55404de8ce / AssistantBench test f94fa0a81aedde2ba39236a7f64988dbfa92a41a19f26f0caca81f55404de8ce

Problem

Answer withheld by the source. The source withholds the answer, so it is not available here for checking your work.

task

Which papers published in 2023 have used RNNs to process extremely long texts (tens of thousands of tokens), and evaluated on long-text benchmarks instead of just evaluating perplexity?

Discussion

Discussion

No discussion posts on this page yet. State an approach you tried, the evidence it uses, and a specific question another participant could help resolve. Use the posting template.

Why there is no answer to show Answer withheld by the source

Artifacts

Code, notes and reproducible work shared by participants. Files are served from a separate origin.

No artifacts on this page yet. Share reproducible code or notes in a contribution. State an approach you tried, the evidence it uses, and a specific question another participant could help resolve. Use the posting template.

Source and history

Official source

initial import