.@llama_index offsite 2026
LlamaIndex held an offsite event in 2026, as announced by Jerry Liu, the project's founder, on social media. The post likely covered updates on the project's direction, team initiatives, or announceme
Knowledge catalogue
LlamaIndex held an offsite event in 2026, as announced by Jerry Liu, the project's founder, on social media. The post likely covered updates on the project's direction, team initiatives, or announceme
Many AI agents in finance rely on extremely high quality context engineering from documents 📑 They can be roughly divided into two categories: 1️⃣ Repetitive, operational work common in back-office us
This post celebrates the arrival of spring by sharing observations or images of young waterfowl, specifically ducklings and goslings, which are iconic symbols of the season. The post is from Yohei Nak
'mystery street meat' he ate a skewer 🤦 Jensen took a last-minute flight to Beijing. Only took one bag and two outfits. Landed for a party, then spent a day eating mystery street meat and noodles whil
This post likely discusses Yohei Nakajima's understanding of stateful AI agents—systems that maintain and utilize information across multiple interactions rather than treating each conversation indepe
Alison Weissbrot / Adweek: Publicis agrees to acquire LiveRamp, which allows companies to share and build new data sets and models that can power agentic frameworks, for 2.2B in cash — Publicis Groupe
Recently @pinecone introduced Nexus – a new knowledge-engine layer for AI agents that reduces token use by up to 90%. It’s built on top of a vector database, but shifts reasoning earlier in the pipeli
symbolic tools like those listed below - rather than pure scaling -are likely driving most of the advances now. once you realize that you realize that hyperscaling is probably insane. @GaryMarcus IMHO
Jerry Liu, likely the creator or leader of LlamaIndex, attended the aiDotEngineer event in Singapore, which featured a workshop, keynote presentation, and executive dinner for developers. The post tha
The best feature of @xai Grok Build right now is how it handles subagents and personas. Most people still treat the model like one very smart intern that has to do everything at once. Grok Build took
This is one of the first real continual learning systems for agents in production. Not just monitoring. Actually getting better over time. LangSmith Engine is how we’re spinning the always-on, self-im
We gave a full 90 minute workshop on how to build agentic workflows over your enterprise documents at @aiDotEngineer Singapore 🇸🇬🦙 The majority of unstructured information is locked up within PDFs. @h
WebWright is a Chrome browser extension that enables AI agents to automate web tasks directly within the browser, leveraging local models like those available through Ollama. The announcement on r/oll
This post likely discusses how developers often create unnecessarily complex evaluation environments when testing AI agents, and suggests simplifying these setups for more effective testing. The advic
Are your benchmarks actually measuring the capability you think they measure? New paper says they probably not. Coined the 'The Evaluation Trap', it provides a vocabulary for auditing whether your eva
Closing out day 2 of @aiDotEngineer Singapore. @swyx a man for the people! 🛑[LIVE] on day 2 of @aiDotEngineer 🇸🇬 @swyx the man himself on building @aiDotEngineer, what true enterprise focus looks like
Singapore's head of AI Govtech estimates the country will have 1.3 billion AI agents deployed within the next two years and is actively developing infrastructure to support this expansion. This sugges
gotta say Codex is completely unrecognizable from 3 months ago. guys went extreme founder mode on this thing @gabrielchua was demoing this and i was like “you guys have agentic excel on mac” @Gavriel_
Grok Build has three commands for managing memory across sessions: /memory, /flush, and /dream. They're experimental but worth looking at if you've ever been frustrated with how agents forget everythi
If you're ascendant/immortal+, we'd love to chat: Green flag if you don't bait your team or mald at them * If you instalock Reyna you're a cracked IC * If you play Brim/Viper you're probably good with
Interesting interpretability paper on tool-using agents. The authors probe hidden states and find the model often recognizes it should call a tool, but fails to actually call one. The mismatch ranges
Dominic-Madori Davis / TechCrunch: Nectar Social, which offers an agentic OS for marketers, raised a 30M Series A led by Menlo Ventures, with GV and True Ventures among investors — AI-powered marketin
Jerry Liu announced that an upcoming team offsite will feature a photo session with a cult aesthetic or theme. This suggests LlamaIndex (the organization Liu is associated with) is planning a creative
the Codex app is in a category of its own. “agentic excel on mac” is an interesting description. gotta say Codex is completely unrecognizable from 3 months ago. guys went extreme founder mode on this
Gary Marcus and Brian Greene discussed at the World Science Festival how scaling neural networks alone does not lead to artificial general intelligence, and that combining scaling with external tool u
We just overhauled our built in Notion skill to take advantage of the new ntn CLI! Check out the docs: https://hermes-agent.nousresearch.com/docs/user-guide/skills/bundled/productivity/productivity-no
We're still celebrating the moms who build, Mother's Day or not. Ruth is a designer who spent years making digital products without ever learning to code. The moment Replit's AI agent landed in the ID
xAI has expanded access to X Premium+ subscribers in Hermes Agent. Enjoy! You can now use X Premium subscriptions in Hermes Agent, and Hermes Agent can now search X posts. https://x.ai/news/grok-herme
You can now use X Premium subscriptions in Hermes Agent, and Hermes Agent can now search X posts. https://x.ai/news/grok-hermes You can now use your @grok subscription inside @NousResearch Hermes Agen
Your agent also will now automatically have X Search tool available when using your Grok subscription as well! See details here: https://hermes-agent.nousresearch.com/docs/user-guide/features/x-search
As AI moves deeper into enterprise operations, the industry is entering a new phase: AI resilience. Autonomous systems and fragmented data are reshaping the modern infrastructure stack, forcing organi
arXiv:2605.13850v1 Announce Type: new Abstract: Existing frameworks for LLM-based agent architectures describe systems from a single perspective: industry guides (Anthropic, Google, LangChain) focus o
arXiv:2605.14266v1 Announce Type: new Abstract: Integration of artificial intelligent (AI) agents in higher education is transforming teaching, learning and administrative processes. Although existing
Agent risk management is evolving rapidly as AI moves into consequential decision-making arenas, collapsing the boundary between human and machine risk across the enterprise. The dual-threat landscape
Yohei Nakajima discusses how artificial intelligence has the potential to serve as a unifying force by bridging divides and fostering connection across different groups and perspectives. The post like
arXiv:2605.15034v1 Announce Type: cross Abstract: Large language models (LLMs) have been extensively studied from computational and cognitive perspectives, yet their behavior as communicative actors i
also thought this was cool from the creative hackathon. extended the dashboard into a frontend dev agent. giving an agent a tight use case and surfacing it via a mini chat right next to the artifact i
arXiv:2605.15132v1 Announce Type: new Abstract: Autonomous multi-agent systems based on large language models (LLMs) have demonstrated remarkable abilities in independently solving complex tasks in a
arXiv:2605.15187v1 Announce Type: new Abstract: A bottleneck in learning to understand articulated 3D objects is the lack of large and diverse datasets. In this paper, we propose to leverage large lan
arXiv:2605.15198v1 Announce Type: cross Abstract: Visual reasoning, often interleaved with intermediate visual states, has emerged as a promising direction in the field. A straightforward approach is
arXiv:2605.14054v1 Announce Type: new Abstract: Achieving robust perception-reasoning synergy is a central goal for advanced Vision-Language Models (VLMs). Recent advancements have pursued this goal v
Beau Rothrock had been at @AngelList for two months when he walked into a Redshift-to-Snowflake migration in deep trouble, already two months behind schedule. He had a 5-week window to migrate all 14,
// Beyond Individual Intelligence // One of the more useful multi-agent surveys I've read this year. 200+ papers mapped along three axes: collaboration mechanisms, failure attribution, and self-evolut
arXiv:2605.14892v1 Announce Type: new Abstract: LLM-based autonomous agents have demonstrated strong capabilities in reasoning, planning, and tool use, yet remain limited when tasks require sustained
arXiv:2510.02837v2 Announce Type: replace Abstract: Although recent tool-augmented benchmarks involve complex requests, evaluation remains limited to answer matching, neglecting critical trajectory as
Three years of enterprise AI investment, and most companies are still waiting for the payoff. Boomi LP thinks it has found it, unveiling Boomi Companion, a collection of open-source agent skills that
arXiv:2508.02332v3 Announce Type: replace Abstract: The performance of Bayesian optimization (BO), a highly sample-efficient method for expensive black-box problems, is critically governed by the sele
BREAKING: The results are in for Slides Arena... @AnthropicAI and @Zai_org models continue to lead the way in soft-verifiable domains 1st: Opus 4.7 by @AnthropicAI 2nd: Opus 4.7 (Thinking) by @Anthrop
Bring Cava here Soulva being so popular on Doordash is a sign of SF's lack of variety when it comes to quick dining options. Their meats are drier than sand, their kale is rigid and bitter. It's such
Learn about the experimental general-purpose accessibility agent that GitHub is piloting. The post Building a general-purpose accessibility agent—and what we learned in the process appeared first on T
arXiv:2605.13918v1 Announce Type: cross Abstract: Automated game testing is important for verifying game functionality, but it remains a costly and time-consuming process. Manual testing often misses
arXiv:2605.15041v1 Announce Type: new Abstract: Tool use extends large language models beyond parametric knowledge, but reliable execution requires balancing appropriate reasoning depth with strict st
arXiv:2605.14102v1 Announce Type: new Abstract: Autonomous language-model agents increasingly combine planning, tool use, document processing, browsing, code execution, and verification loops. These c
arXiv:2605.14911v1 Announce Type: new Abstract: High-fidelity physics simulation is essential for closing the sim-to-real gap in robotics and complex mechanical systems. However, the computational ove
arXiv:2601.15620v2 Announce Type: replace Abstract: The 1-identification problem is a fundamental pure-exploration problem in multi-armed bandits. An agent aims to determine whether there exists an ar
arXiv:2605.14398v1 Announce Type: new Abstract: World models have emerged as a powerful paradigm for building interactive simulation environments, with recent video-based approaches demonstrating impr
arXiv:2605.15077v1 Announce Type: cross Abstract: Function calling, also known as tool use, is a core capability of modern LLM agents but is typically constrained by synchronous execution semantics. U
arXiv:2605.14495v1 Announce Type: cross Abstract: Multimedia verification requires not only accurate conclusions but also transparent and contestable reasoning. We propose a contestable multi-agent fr
arXiv:2605.15016v1 Announce Type: cross Abstract: As large language models empower healthcare, intelligent clinical decision support has developed rapidly. Longitudinal electronic health records (EHR)
This post shares feature requests and ideas inspired by experimentation with micro (likely a small language model or tool), credited to another developer. It appears to be a wishlist of desired capabi