Models
Grok Imagine Video 1.5 Lite has a price of $0.14 for each second
Artificial Analysis writes that it makes a 10 second 1080p video in a median time of 60.5 seconds, the lowest time at its quality.
What matters in AI.
SubscribeNews of the day
40 stories filed under this day.
Models
Artificial Analysis writes that it makes a 10 second 1080p video in a median time of 60.5 seconds, the lowest time at its quality.
Security
A model writes each report and no person examines it, and Anthropic writes that some reports can be incorrect.
Industry
Google wants to pay $10 million for about 100 million emails, but the lawmakers write that removal of names does not always make data anonymous.
Security
A developer can select the Locked Down level to stop an agent from the use of the internet and files.
Research
Justin Drake writes that AI mathematics could break wallet signatures, but Buterin writes that it is not necessary to move funds today.
cointelegraph.comClaimed, not confirmed
Industry
One task is $1.77 for this model and $11.67 for Grok 4.7 xhigh.
Agents
A user can click a number to see the query that made it, and can move the dashboard to tools such as Grafana or Hex.
Industry
Jefferies writes that access to websites is the primary problem, and it can slow the revenue that Meta wants.
Benchmarks
Artificial Analysis says a hallucination is material if it can mislead a reader, for example with an incorrect contract date.
Security
Incognia writes that a legitimate agent does not show that the customer wants the task.
Security
Researchers say a user must disable the token monitor of the worm before the user revokes tokens.
Models
The price at 1K resolution is half the price of Nano Banana 2.
Security
In tests, the monitor found 94% of malicious hacking sessions but sent 8.7% of safe sessions for a check.
Agents
The White House has a plan to let users do tasks, such as passport renewal, with AI agents before the end of 2026.
Industry
Most of the 94 US reactors got an AI tool for their work.
Security
The confirmation is on one path, and the attack used a different path to the same result.
straiker.aiClaimed, not confirmed
Benchmarks
The author tells that the correct test was more work than the fast code.
towardsdatascience.comClaimed, not confirmed
Security
Cogent tells that its study found 3 new attack paths for AI agents for each 1 for human attackers.
Agents
Crossmint tells that the toolkit has 1 API for the work of 4 or more vendors.
Security
ASOS tells that the attacker did not get payment information.
Security
AWS made changes for part of the problem, and Zenity tells companies to give each agent a role with only the access it must have.
the-decoder.comClaimed, not confirmed
Policy
The ICO also wrote to OpenAI, Anthropic and Meta about AI agents that did not obey their safeguards.
Security
AWS corrected the default permissions of the agents between June 22 and September 29, Zenity wrote.
www.csoonline.comClaimed, not confirmed
Security
Upgrade the SDK to version 2.2.0 or 1.30.0 to remove the vulnerability.
Security
The text works in only about 35% of tests at most, and it is always plaintext, which tools can find.
Research
The authors tell that 2 hosted call sites give the same live result for 28% of the spend.
arxiv.orgClaimed, not confirmed
Security
The attacks made the data of 68,000 or more customers open to the attackers.
The SOFX ReportClaimed, not confirmed
Agents
The authors tell that, with the gate, 95% of tasks are correct, not 65%.
arxiv.orgClaimed, not confirmed
Security
The authors tell that the monitor also gave an alarm for 29.3% of safe runs.
arxiv.orgClaimed, not confirmed
Security
The authors put malicious prompts in data that looks safe, and the attack works only after a world model uses the data.
arxiv.orgClaimed, not confirmed
Security
The authors found a false-positive rate of 0.1%, and the system limits the model to a closed set of commands.
arxiv.orgClaimed, not confirmed
Models
In tests by Cloudflare, the 9B Clef-Flash model has a median latency of 38.8 ms, against 209.3 ms for the 27B Clef model.
Research
The authors found that humans, DeepSeek and Qwen3-Max all cut standard delivery by 48 to 53 percentage points. The results are different for each value of time.
Bioengineer.orgClaimed, not confirmed
Research
The authors found 12.3 times more collisions in real-time tests than in static tests. The agents completed 91% to 94% of the tasks.
arxiv.orgClaimed, not confirmed
Research
The authors found a typical error of 2.9 times for Fable 5.1 in Claude Code, but only 1.2 times for GPT-6 Astra in Codex.
arxiv.orgClaimed, not confirmed
Security
The researchers find a success rate of more than 95% for the attack.
arxiv.orgClaimed, not confirmed
Research
2 easy methods to add diversity did not decrease the risk.
arxiv.orgClaimed, not confirmed
Research
The researchers find that this reporting destroys 68% of the gains from delegation.
arxiv.orgClaimed, not confirmed
Research
An instruction to randomize independently decreased the correlation between agents with the same input, but did not remove it.
arxiv.orgClaimed, not confirmed
Security
The attack can also get around 3 agent-level defenses.
arxiv.orgClaimed, not confirmed