Research
A 7M-parameter model gets 97.9% exact accuracy on Sudoku-Extreme
In InfiLoop, accuracy continues to increase after more than 20,000 effective steps.
arxiv.orgClaimed, not confirmed
What matters in AI.
SubscribeTag
47 stories carry this tag.
Research
In InfiLoop, accuracy continues to increase after more than 20,000 effective steps.
arxiv.orgClaimed, not confirmed
Research
The adapters gave 84.9% top-1 precision on new documents, against 63.4% for the model with no adapter.
arxiv.orgClaimed, not confirmed
Compute
The framework keeps 98.5% of the reference mAP and works on different GPU backends.
arxiv.orgClaimed, not confirmed
Agents
At each step, it follows the step in memory, changes its parameters, or gives the step to the base agent.
arxiv.orgClaimed, not confirmed
Research
It puts adapters only in layers with low input-output cosine similarity, and it increases the average target-task performance.
arxiv.orgClaimed, not confirmed
Agents
On three datasets, its Recall at 1 increases by a maximum of 16.18%.
arxiv.orgClaimed, not confirmed
Agents
The authors say that harness changes repair process failure and weight training repairs content failure.
arxiv.orgClaimed, not confirmed
Research
The authors say an LLM with no such step did not end the deadlock in 5 tests.
arxiv.orgClaimed, not confirmed
Research
The authors say 14.6% to 49.7% of the fixes removed the bug, by agent.
arxiv.orgClaimed, not confirmed
Research
The authors say it sent 18.4% of the 1,247 test alerts to analysts.
arxiv.orgClaimed, not confirmed
Research
A student model with fine-tuning on number data hacks in 58.3% of chess episodes, against 10.9% with no fine-tuning.
arxiv.orgClaimed, not confirmed
Benchmarks
The researchers find that the resolution rate of the agents is 33.2% and not 50.6% when TestJack does the check.
arxiv.orgClaimed, not confirmed