Amazon Nova Act is now HIPAA eligible
Amazon Nova Act is a service for building and managing AI agents that automate browser-based UI workflows, powered by a custom Nova 2 Lite model. It is particularly valuable for industries with fragme
Knowledge catalogue
Amazon Nova Act is a service for building and managing AI agents that automate browser-based UI workflows, powered by a custom Nova 2 Lite model. It is particularly valuable for industries with fragme
As of yesterday, we made everyone within @llama_index research/engineering/product a Member of Technical Staff. Yes all the frontier labs have done this for years. But we didn’t only make these change
arXiv:2605.20811v1 Announce Type: new Abstract: Robotic imitation learning is often treated as reproducing demonstrated actions, but actions are inherently embodiment-specific. When demonstrations com
Partner ecosystems and hybrid AI orchestration have become essential for enterprises navigating agentic AI at scale, where governance, data sovereignty and coordination determine success in a regulate
🚀Qwen3.7-Max just landed at 56.6 on the Artificial Analysis Intelligence Index — a solid 4.8pt jump over Qwen3.6-Max-Preview. @ArtificialAnlys ⚡️Sharper sci reasoning, stronger agentic chops, better c
Tips for using Grok Build If you are managing multiple machines across clusters and farms, you could ask Grok Build to spin off a sub-agent SSH tunnel into the machines while interacting with tmux ses
Cybersecurity and password service provider 1Password LLC today expanded its collaboration with OpenAI Group PBC, releasing a Model Context Protocol server that lets the Codex coding agent pull creden
deepagents v0.6 ships w/ support for code interpreters! these are the perfect happy medium between pure tool execution and heavyweight sandboxes they give your agent an environment where it can... → k
arXiv:2605.18881v1 Announce Type: new Abstract: In dynamic flow fields, various animals exhibit remarkable odor search capabilities despite relying on stochastic detections. Interestingly, there exist
Hybrid AI governance has become essential for regulated industries such as banking, where the pressure to move fast with agentic AI collides with strict requirements for data sovereignty, compliance a
arXiv:2605.18772v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) improves the factual accuracy of large language model (LLM) outputs by grounding generation in external knowledge
arXiv:2605.20072v1 Announce Type: new Abstract: Large Language Models are increasingly proposed as cognitive components for robotic systems, yet their opaque decision processes make it difficult to ex
arXiv:2605.18784v1 Announce Type: cross Abstract: The rapid diffusion of agentic AI has created a new coverage problem for commercial insurance: some AI-mediated losses are now affirmatively insured,
We are also launching Science Skills, a specialized bundle that integrates insights from 30+ major life science models and databases with agentic platforms like @Antigravity to allow researchers to pe
arXiv:2605.18133v1 Announce Type: cross Abstract: LLM-based chatbot agents increasingly process user requests by combining natural-language reasoning with external tools such as web browsing. These ca
arXiv:2601.21841v3 Announce Type: replace Abstract: While Large Language Models (LLMs) have demonstrated strong zero-shot reasoning capabilities, their deployment as embodied agents still faces fundam
In this post, we demonstrate how you can extend the conversational memory of Kiro CLI by implementing a custom Model Context Protocol (MCP) server that integrates with Amazon Bedrock AgentCore Memory.
arXiv:2605.17812v1 Announce Type: new Abstract: Vertical AI firms in accounting, law, healthcare, procurement, and similar domains historically bundled workflow, domain logic, and accountability into
Google LLC’s Android team today announced it’s employing deeper integration of generative artificial intelligence and agentic AI to help mobile developers and publishers connect with customers and use
Google is making a big push into cybersecurity. At I/O, the company announced that it was inviting select groups of experts to test the API for CodeMender, an 'AI agent for code security' it debuted l
arXiv:2605.17426v1 Announce Type: cross Abstract: We propose a framework for predicting the effects of mobility introduction measures using a human-flow digital twin. This digital twin incorporates a
arXiv:2508.00712v2 Announce Type: replace-cross Abstract: We introduce JSON Bag-of-Tokens model (JSON-Bag) as a method to generically represent game trajectories by tokenizing their JSON descriptions
arXiv:2605.17907v1 Announce Type: cross Abstract: By sharing intermediate features, collaborative perception extends each agent's sensing beyond standalone limits, but real-world feature modality hete
arXiv:2605.16727v1 Announce Type: new Abstract: We introduce PopuLoRA, a population-based asymmetric self-play framework for reinforcement learning with verifiable rewards (RLVR) post-training of LLMs
arXiv:2605.16407v1 Announce Type: cross Abstract: We present a framework for verifying the deterministic structured computations surrounding a large language model rather than the model itself, extend
arXiv:2605.16530v1 Announce Type: new Abstract: Realistic surgical simulation plays a crucial role in training novice surgeons and in the development of autonomous agents. World models can scale such
The daddy of baby AGI articulates what we all feel is missing in the current state of AI architecture I didn’t really understand what people meant by stateful agents so I started exploring my current
arXiv:2605.17064v1 Announce Type: new Abstract: Large language models optimized for instruction following and agentic tasks remain poorly aligned with the requirements of high-quality creative writing
The rise of intelligent digital workers and autonomous AI agents is compounding an already urgent cybersecurity challenge: attack surfaces expanding faster than security teams can manually assess them
In this post, you will implement four Lambda-based custom code evaluators for a financial market-intelligence agent, register each with AgentCore, and run them in on-demand and online modes. You will
arXiv:2605.15938v1 Announce Type: cross Abstract: Finding an odor source in a turbulent flow requires effectively leveraging the history of olfactory observations into a robust navigation strategy. In
arXiv:2605.15225v1 Announce Type: cross Abstract: Biologically-inspired AI agent frameworks claim reliability benefits through structural guarantees adapted from gene regulatory networks, immune syste
The rise of agentic AI is reshaping what enterprise partnerships must deliver — and domain-specific AI is proving to be the real differentiator between systems that advise and systems that act. As AI
The AI governance gap is widening rapidly as AI agents assume operational roles inside enterprise workflows, expanding the security perimeter faster than most organizations can mitigate exposure. But
Introducing http://Browse.sh, the largest open-source catalog of skills to reliably perform any task on the internet. We've researched hundreds of sites to give your agents the playbook they need to n
Dust, an agentic artificial intelligence startup that’s trying to push enterprise workers away from isolated chatbots into a more collaborative, multiplayer ecosystem, said today it has raised 40 mill
arXiv:2605.15227v1 Announce Type: new Abstract: Self-driving laboratories (SDLs) have attracted increasing attention as a means of accelerating scientific discovery; however, developing SDL software r
Try it out … Improvements are landing every few days! Grok Build CLI Beta can now be installed directly from Grok Web with a single terminal command. The agentic coding and workflow tool is currently
Woke up to see that Grok Build finished my feature build from last night. But what's the most interesting to me, is that it has a set of suggestions for what to do to really make sure everything is do
gotta say Codex is completely unrecognizable from 3 months ago. guys went extreme founder mode on this thing @gabrielchua was demoing this and i was like “you guys have agentic excel on mac” @Gavriel_
We're still celebrating the moms who build, Mother's Day or not. Ruth is a designer who spent years making digital products without ever learning to code. The moment Replit's AI agent landed in the ID
arXiv:2605.15077v1 Announce Type: cross Abstract: Function calling, also known as tool use, is a core capability of modern LLM agents but is typically constrained by synchronous execution semantics. U
arXiv:2605.14237v1 Announce Type: new Abstract: Deploying AI agents for repetitive periodic tasks exposes a critical tension: Large Language Models (LLMs) offer unmatched flexibility in tool orchestra
arXiv:2603.07833v2 Announce Type: replace-cross Abstract: Temporal-difference (TD) learning is highly effective at controlling and evaluating an agent's long-term outcomes. Most approaches in this par
arXiv:2605.14038v1 Announce Type: new Abstract: Large language models (LLMs) increasingly act as autonomous agents that must decide when to answer directly vs. when to invoke external tools. Prior wor
arXiv:2605.14111v1 Announce Type: new Abstract: Hospital pharmacists make high-stakes decisions to mitigate drug shortages under uncertainty, time pressure, and patient risk. Interviews revealed that
SmithDB lets you see traces in seconds instead of minutes. This quote from our friends at @cogent_security says a lot: “At Cogent, our background agents can produce a huge volume of traces all at once
turns out that building evals is super super challenging even now. i thought a lot of it was table stakes but turns out it has only become harder since agents are now more complex than ever! going to
Agentic AI might be outrunning the one thing it needs most — an AI trust layer. That challenge is reshaping the ambitions of companies long anchored in data protection. Veeam Software Group GmbH, whic
arXiv:2512.06471v2 Announce Type: replace-cross Abstract: Goal-conditioned reinforcement learning (RL) concerns the problem of training an agent to maximize the probability of reaching target goal sta
arXiv:2605.06390v2 Announce Type: replace Abstract: A leading proposal for aligning artificial superintelligence (ASI) is to use AI agents to automate an increasing fraction of alignment research as c
Early beta and we are looking forward to your feedback! We will keep improving until Grok Build is absolutely the best in the world An early beta of Grok Build, an agentic CLI for coding, building app
arXiv:2602.05242v1 Announce Type: cross Abstract: Agentic Test-Time Scaling (TTS) has delivered state-of-the-art (SOTA) performance on complex software engineering tasks such as code generation and bu
arXiv:2605.13193v1 Announce Type: new Abstract: Fine-grained recognition in everyday life is often not a closed-book classification problem: when encountering unfamiliar objects, humans actively searc
i recently left @Bloomberg after two great years. incredibly lucky to have joined @LangChain in time to work on this 🚀 Spend less time on triaging Ship fixes faster Catch regressions earlier Introduci
I setup Hermes this week. A lot of you said you want to try it out, I'm documenting the whole journey, week by week, what it learns, what it builds for itself, what this agent actually becomes after a
Install. Try. Share Feedback. Ship stuff. An early beta of Grok Build, an agentic CLI for coding, building apps, and automating workflows is now available for SuperGrok Heavy subscribers. Through this
This Mitchell Hashimoto quote about Bun migrating from Zig to Rust reminded me of a similar conversation I had at a conference last week. I was talking to someone who worked for a medium sized technol
arXiv:2602.10032v2 Announce Type: replace Abstract: Agents in cyber-physical systems are increasingly entrusted with safety-critical tasks. Ensuring safety of these agents often requires localizing th
SmithDB is the perfect example of how far performance can be pushed by having full control over the storage layer. The DataFusion + @vortexdotdev stack seems to be emerging as THE way to build next ge