01PolicyProduct8 sources agree
Anthropic has introduced machine-readable watermarks for text and images generated by its Claude AI models to comply with the EU's AI Act, with the watermarks being deployed worldwide in all models launched after August 2, 2026. The company claims the watermarks will not decrease output quality or be visible to readers, but critics argue they may be ineffective, harmful, or a privacy violation. The move is seen as a significant development in the regulation of AI-generated content, with other companies likely to follow suit.
Provenance — who else covered this
02ModelsProduct2 sources agree
SpaceXAI introduced Grok 4.6, a vision-language model developed with Cursor, offering improved performance and lower prices, and available via API, Grok Build, and Cursor. The model achieved top scores on several benchmarks, including GPQA Diamond and AA-Briefcase, and rivals Claude Opus 5 and GPT-5.6 Sol at a lower cost per task. Grok 4.6 is the result of a partnership between SpaceXAI and Cursor, which led to an acquisition and the development of new models, including Origin, a code hosting service. The model's ability to complete long-running work with fewer turns makes it a significant advancement in the field.
Provenance — who else covered this
03On-deviceInternals7 sources agree
UC Berkeley has open-sourced FreeToken, which achieves 2-4x faster local LLM inference than Ollama, with notable performance on various models and GPU configurations.
Provenance — who else covered this
04AgentsProduct4 sources agree
Researchers at Shanghai Jiao Tong University and others devised a workflow called Agentic ASR, which pairs an automatic speech recognition engine with a large language model to detect and correct transcription errors interactively. The system treats transcription as a multi-turn process of refinement, allowing users to dictate and then confirm or correct the transcription in further turns, and it significantly improved semantic error rates on multilingual speech-to-text benchmarks. The approach has potential applications beyond speech-to-text, such as editing documents, reviewing code, and iterating on a design.
Provenance — who else covered this
05SafetyInternals4 sources agree
A new paper presents a method to reduce reward hacking using multi-agent debate with online RL training, addressing safety risks when models learn to exploit human judges. The approach helps mitigate reward hacking with a weak judge.
Provenance — who else covered this
06RoboticsProduct3 sources agree
Researchers led by Dantong Niu release T-Rex, a tactile-reactive dexterous manipulation dataset, and make it available on Hugging Face, the dataset is intended for use in robotic manipulation tasks.
Provenance — who else covered this
07BusinessProductsingle source
GPT-5.6 Sol API prices are being reduced by over 20% for the next 3 months, with credits going further in Codex on token-based plans, and usage included in subscriptions remaining the same.
Provenance — who else covered this
08AI securityProduct2 sources agree
A new course, Attacking AI, educates on assessing AI-enabled web-apps, APIs, infrastructure, and productivity features for vulnerabilities.
Provenance — who else covered this