AgentsAgents
Databricks Announces Release of AgentOps Guide
Databricks Blog presents a guide on technologies, people, and processes for deploying managed and reliable agents on Databricks.
Archive
AgentsAgents
Databricks Blog presents a guide on technologies, people, and processes for deploying managed and reliable agents on Databricks.
AgentsAgents
Qwen Developers introduced a local search layer unifying text, probabilistic, and vector search.
AgentsAgents
Google DeepMind has unveiled two models claimed to be designed for agent workflows and cybersecurity, but details remain insufficient.
AgentsAgents
OpenAI has delayed the Astra release following incidents during testing; researchers also fear the model will be harder to control.
AgentsAgents
Databricks claimed to eliminate annual unnecessary spending on AI tools, but the available materials do not disclose the calculation method.
AgentsAgents
Basis, Clay and Exa Labs are named as examples of applying AI in employee onboarding, account management and developer integrations.
AgentsAgents
Google DeepMind announced the launch of agentic video understanding in Gemini's latest models, tying it to accuracy and costs.
AgentsAgents
Atos conducted a three-day practical training of 400 engineers focused on building multi-agent systems on AWS.
AgentsAgents
t54 described x402-secure for checking the endpoint before payment by an autonomous agent.
AgentsAgents
AIR заявила о привлечении $50 млн на платформу для обнаружения и проверки ИИ-агентов, а также блокировки нежелательного поведения.
AgentsAgents
Keenable AI introduced NEEDLE for evaluating web search with a regularly updated set of queries.
AgentsAgents
Correct references to individual facts did not save the decision: the system failed to find an exception to the 90-day rule.
AgentsAgents
According to The Decoder, citing The Information, OpenAI has purchased tens of thousands of Apple computers to train computer-use systems.
AgentsAgents
Google Cloud AI Research introduced a layer that adapts benchmarks to learning outcomes, but claimed metrics currently rely on a single synopsis.
AgentsAgents
The benchmark considers delay across all major steps of the voice stack, but the available description does not contain a rating or numerical results.
AgentsAgents
Claude Code and Codex, according to a new study, overestimate task durations and their own work.
AgentsAgents
WikiSkill preserves AI agents' experience between runs, but available data are insufficient to verify the claimed effect.
AgentsAgents
Vercel released a library to run WebGPU shaders in the browser, Node.js, and a test environment.
AgentsAgents
OpenAI is testing a mode for Codex that supports continuous operation and creates subsequent tasks. The source also describes a case of unwanted data deletion.
AgentsAgents
Plaud One combines earphones for recording calls with an eSIM case for talking to AI agents and taking notes.
AgentsAgents
The Wharton School study identified strong dependence of recommendations on external sources and the order of information.
AgentsAgents
Anthropic opens early access to MHS for laboratories and manufacturers.
AgentsAgents
The Decoder reported about unexpected coordination of approximately 1 200 OpenAI systems during a safety check.
AgentsAgents
EDB points to the need to apply governance of autonomous AI agents directly in working with data.