Google proposes contextual privacy checks for AI agents

A new workshop report examines how agents should handle changing tasks and data sharing.
Accessibility Adjustments
Use these optional tools to adjust reading and display preferences. These tools cannot resolve every accessibility barrier. Please contact the website owner if you need assistance.

A new workshop report examines how agents should handle changing tasks and data sharing.

Liquid AI has enabled image inputs in its paid d1 API, expanding structured decisions to visual inspection and screenshot tasks.
Mistral’s new flagship has 1.05 trillion parameters and a million token context window. Downloadable weights are planned for later in October.

The research preview combines visual generation and understanding within shared context.

A September check of three synthetic images found uneven use of embedded provenance, with important limits on what the results establish.

The evaluator tests security boundaries before measuring a model’s cyber capabilities.

Its employer data shows rising mentions of applied AI skills and a decline in foundational skills from their peak.

The download combines a reasoning core with specialist tools and needs a GPU with 80 GB of memory.

Early access is limited while Reflection completes safety evaluations.

A robotics study tests direct video pretraining before labeled robot adaptation

A distillation study measures the tradeoff between robot policy speed and success

The October 5 funding announcement adds the seed amount and lead investors. The robot testing business remains available through private access.

Human video training improves a manipulation model in simulated tests

A Keio study tests a shared motion representation across robot demonstrations

More flexible delegation raised benchmark scores and the cost per task.

Hello Robot outlines a new research phase with home studies and care staff evaluations

Memory checks reduce attacks in a controlled agent study

Runtime selection leads a small AI scientist comparison

Graphite compares Opus 5.5 with human writing on 9,974 topics

ThinkingBox compares average completion and repeated success

Microsoft’s evaluation separates trace linkage from reliable agent diagnosis

Lightroom desktop 9.6 brings early access editing from text prompts, separate DNG outputs and generative credit costs

Robot recovery research reports gains across four physical tasks

VeriSpec checks written AI rules and leaves final judgments to reviewers