love this framing of memory as 'proactive'
love this framing of memory as 'proactive' agent memory has always been reactive. OpenWiki makes it proactive. connect to sources, tell it what you care about, and your agent hits the ground running .
Knowledge catalogue
love this framing of memory as 'proactive' agent memory has always been reactive. OpenWiki makes it proactive. connect to sources, tell it what you care about, and your agent hits the ground running .
Saritha Rai / Bloomberg: Malaysia Prime Minister Anwar Ibrahim plans to debut an agentic AI avatar of himself within days, which is meant to help the public navigate government services — Malaysia Pri
Many people were doubting Meta's position in the AI race. Yesterday, they dropped Muse Spark 1.1, now one of the strongest agentic models, and massively undercut OpenAI and Anthropic on price. When I
arXiv:2607.08080v1 Announce Type: new Abstract: Aspect Sentiment Triplet Extraction (ASTE) requires jointly identifying (aspect, opinion, sentiment) triples from a given review sentence. While large l
arXiv:2607.07857v1 Announce Type: cross Abstract: We build a team of specialized large language-model agents and present an agent-driven workflow for research-level formalization in theoretical physic
arXiv:2607.08282v1 Announce Type: cross Abstract: While Large Language Models (LLMs) have become essential productivity tools, their integration into workflows without adequate safeguards creates sign
arXiv:2607.08647v1 Announce Type: cross Abstract: As autonomous agents are increasingly deployed across diverse operational contexts, aligning their behavior with human intent demands reward functions
arXiv:2607.08391v1 Announce Type: cross Abstract: Making tradeoffs between execution latency and result utility (i.e., anytime computing) for adapting to dynamic operational requirements has been show
OpenSWE is one of our most widely used agents throughout the company. Since July 1st it's been tagged over 700 times in Slack! This doesn't even count the reviewer agent, tagging it in GitHub, or tagg
arXiv:2607.08180v1 Announce Type: cross Abstract: The rise of LLM-based agents with reasoning, summarization, and memory capabilities has created a new threat surface for online content that conventio
arXiv:2607.08233v1 Announce Type: new Abstract: A central challenge in building intelligent systems is enabling agents to jointly perceive complex inputs, form hypotheses about hidden patterns, and de
arXiv:2607.08691v1 Announce Type: cross Abstract: Repository-level code generation requires implementing target functions while accounting for complex cross-file dependencies and project-specific conv
arXiv:2607.08012v1 Announce Type: cross Abstract: This paper studies an online variant of the assistance games framework, where an informed agent and an uninformed agent repeatedly interact over T tim
In this post, we show you how to combine case management with agentic automation capabilities in Quick Automate. We introduce case management and explore the lifecycle of cases in an agentic workflow
Cursor has released updates featuring new cloud agent hooks and side-chat functionality, expanding its AI-assisted development capabilities. The changelog details recent features and improvements to t
arXiv:2607.08373v1 Announce Type: cross Abstract: Connected vehicles are autonomous cyber-physical systems whose behavior must be continuously monitored during operation to detect deviations from norm
Sierra isn't the first to build this - Ramp, Stripe, CoinBase also have If you want an open source version - check out OpenSWE: https://github.com/langchain-ai/open-swe We use it internally (mostly fo
arXiv:2601.02871v3 Announce Type: replace Abstract: Task-oriented proactive dialogue agents play a pivotal role in recruitment, particularly for steering conversations towards specific business outcom
arXiv:2503.08936v3 Announce Type: replace-cross Abstract: Scenario-based testing with driving simulators is extensively used to identify failing conditions of automated driving assistance systems (ADA
arXiv:2607.08565v1 Announce Type: cross Abstract: LLM scheduling is critical to serving, yet it remains unclear how well existing designs fit agentic serving--with LLM requests issued by agents instea
Zijing Wu / Financial Times: Sources: Manus investors and management are discussing unwinding Meta's $2B buyout at the same valuation, with Tencent in talks to become the largest investor — Chinese te
arXiv:2607.07873v1 Announce Type: new Abstract: The scalability of organic agriculture is partially limited by the labor costs associated with monitoring for pests. While drones and rovers are well-su
arXiv:2607.08402v1 Announce Type: cross Abstract: Large-scale and diverse datasets are needed to train AI models to take real-time decisions for autonomous vehicles (AVs), an intelligent transportatio
Agentic marketing leverages AI agents to automate marketing tasks, and this approach requires a robust data foundation as its core component. Databricks discusses how organizations need unified data p
arXiv:2607.08495v1 Announce Type: cross Abstract: Sharp et al. (2025) introduce 'agentic inequality' as a framework for analyzing disparities in access to AI agents across three dimensions: availabili
The most important thing about Grok Build and the 4.5 release is that it is genuinely so useful for real-world work Grok 4.5 just topped Perplexity’s WANDR orchestrator evaluation It scored higher tha
this is a great question! I think OpenWiki and memory wikis in general can be thought of as a type of knowledge graph. They’re different from traditional knowledge graphs in the sense that these are j
arXiv:2607.07885v1 Announce Type: cross Abstract: Dynamic obstacle avoidance in unstructured outdoor environments remains a critical challenge for autonomous mobile robots, particularly when large-sca
arXiv:2607.08010v1 Announce Type: new Abstract: Production LLM agents often waste latency and reliability by regenerating code for the same procedural steps on every request. We replace this inference
arXiv:2607.08124v1 Announce Type: cross Abstract: The behavior of an LLM agent is determined not only by the underlying model, but also by its harness: the executable program that constructs context,
We’re Hiring: Software Engineer (R&D, Infrastructure and Platform Reliability) 🐟 https://sakana.ai/careers/infrastructure-platform-reliability-engineer/ Sakana Fugu, our Multi-Agent System as a Model,
The AI engineering world is using “loop” to describe several different agent architectures. This post maps execution loops, task loops, product loops, system loops, and the human oversight loop that c
arXiv:2607.07847v1 Announce Type: new Abstract: As large language models (LLMs) become increasingly capable, the next question is how can we enable models to continually learn? Today, the field largel
arXiv:2607.06990v1 Announce Type: new Abstract: Multi-robot systems provide the parallelism and redundancy necessary for long-horizon tasks, while Large Language Models (LLMs) offer the reasoning capa
A self-improving agent is conceptualized as a software factory—a system capable of generating and modifying code—that is directed toward improving its own codebase and capabilities. This framing sugge
arXiv:2607.07475v1 Announce Type: new Abstract: In robotics, the capability of an artificial agent to represent the range of its action possibilities, i.e. affordances, is crucial to understand how it
Infrastructure design is being redefined by agentic AI, pushing the industry toward system-level AI infrastructure optimization, balancing performance and cost across diverse workloads rather than foc
big update to openwiki - it now supports two modes: - code brain: create a wiki for a code base - personal brain: create a wiki for more general purpose tasks (email, web, etc) webinar at 11am PST (5
Built from scratch by Grok 4.5 + Grok Build in UE5.8: a cyberpunk L-corner street with neon facades, rain, signs, and crowds walking through the scene. End-to-end, the run took 10.75M tokens, 36.5 min
This OpenAI announcement highlights ChatGPT's expanded capabilities and positioning as a tool for complex, professional work rather than just simple queries. It likely covers new features, improved pe
arXiv:2607.06700v1 Announce Type: new Abstract: Multi-agent Simultaneous Localization and Mapping (SLAM) and collaborative SLAM (CSLAM) require robots to continuously exchange global descriptors (GDs)
arXiv:2607.06691v1 Announce Type: new Abstract: Human-human collaboration is a fundamental aspect of everyday life, essential to success in a wide range of goal-directed activities from household task
arXiv:2508.00554v4 Announce Type: replace-cross Abstract: In financial trading, large language model (LLM)-based agents demonstrate significant potential, but their decisions can be sensitive to noisy
As agentic AI accelerates enterprise transformation, data sovereignty is crystallizing from a compliance checkbox into a foundational strategic imperative — one that determines not just where data liv
arXiv:2511.23347v2 Announce Type: replace Abstract: An associative memory (AM) enables cue-response recall, and it has recently been recognized as a key mechanism underlying modern neural architecture
arXiv:2607.07139v1 Announce Type: new Abstract: Underwater robots often operate near delicate targets where high-power thrusters resuspend sediments and induce turbulence, degrading image quality at t
arXiv:2603.21489v2 Announce Type: replace-cross Abstract: AI agents have become increasingly capable at isolated software engineering (SWE) tasks such as resolving issues on Github. Yet long-horizon t
arXiv:2607.06964v1 Announce Type: cross Abstract: Bridging the gap between human pilot intent and autonomous flight operation is critical for real-world electric vertical takeoff and landing (eVTOL) a
For coding agents, trustworthiness has to be tested in the harness where the model actually writes code. In one surveillance scenario, Kimi K2.7 complied with the request in 8/8 samples. SWE-1.7 refus
arXiv:2607.06786v1 Announce Type: cross Abstract: Standards bodies, including TM Forum, 3GPP, and ETSI, are converging on Agentic AI as the foundation for next-generation network management, where Lar
arXiv:2607.07321v1 Announce Type: new Abstract: Tool utilization enables Large Language Model (LLM) agents to interact with the real world and resolve complex tasks. However, existing agent frameworks
arXiv:2607.07702v1 Announce Type: new Abstract: The optimization of long-horizon agents increasingly relies on reflection-based mechanisms, where a large language model (LLM) acts as an optimizer to d
Fun conversation with @swyx on our journey building the cloud for true elastic inference, sandboxes, and more. And of course, how we're evolving Modal's dev experience to be better for agents. Modal's
arXiv:2607.07626v1 Announce Type: cross Abstract: Reliable confidence estimation is essential for deploying large language models (LLMs) in confidence-aware systems, where downstream decisions such as
arXiv:2607.07029v1 Announce Type: cross Abstract: Reinforcement learning (RL) policies can be unsafe and vulnerable to attacks. Ensuring their reliability is often a pain point as existing automated t
Grok 4.5 on OpenClaw Grok 4.5 from @SpaceXAI is live on OpenClaw. No OpenClaw update required, just connect your X Premium or SuperGrok subscription, select Grok 4.5 under the xAI provider, and use an
Yohei Nakajima created a virtual dance studio application in Replit that uses computer vision to analyze uploaded dance videos and automatically generates corresponding 3D rig animations. The project
arXiv:2607.06619v1 Announce Type: cross Abstract: Modern processor verification struggles to reach deep architectural states due to the inefficiencies of traditional mutation-based fuzzing. We propose
A demonstration of an AI agent with iPhone integration that autonomously ordered a donut, showcasing early progress in agentic AI capabilities. The demo was presented publicly by AGI Inc and Div Garg,
Hermes Agent is a project from Nous Research focused on developing agentic AI systems capable of autonomous reasoning and task execution. The initiative likely explores methods for creating AI agents