AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,964 results
Agents

Reflection in the Dark: Exposing and Escaping the Black Box in Reflective Prompt Optimization

DGX agent

arXiv:2603.18388v2 Announce Type: replace Abstract: Automatic prompt optimization (APO) has emerged as a powerful paradigm for improving LLM performance without manual prompt engineering. Reflective A

agentsarxiv-cs-ai
9 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

ResearchClawBench: A Benchmark for End-to-End Autonomous Scientific Research

DGX agent

arXiv:2606.07591v1 Announce Type: cross Abstract: AI coding agents are increasingly used for scientific work, but their end-to-end autonomous research capability remains difficult to verify. We presen

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Scaffold Effects on GAIA: A Controlled Comparison

DGX agent

arXiv:2606.08529v1 Announce Type: new Abstract: Published agent capability scores conflate what a model can do with what its scaffold lets it do, and the magnitude of this elicitation gap is not well

model-releasesarxiv-cs-ai
9 Jun 2026
Agents

Trustworthy Smart Fabs via Professional Proxies: Scaling Safe and Sustainable by Design (SSbD) through Industrial Data Spaces

DGX agent

arXiv:2606.09227v1 Announce Type: cross Abstract: The convergence of the 2026 European Union Safe and Sustainable by Design (SSbD) framework, Corporate Sustainability Due Diligence Directive (CSDDD),

agentsarxiv-cs-ai
9 Jun 2026
Safety

Accounting for Context: Shaping Moral Credences for Value Alignment

DGX agent

arXiv:2606.06972v1 Announce Type: new Abstract: Ensuring that agent behaviours are aligned with human moral values inevitably raises the problem of how to account for the plurality of moral perspectiv

safetyarxiv-cs-ai
8 Jun 2026
Agents

SCOUT: Semantic scene COverage via Uncertainty-guided Traversal

DGX agent

arXiv:2606.06721v1 Announce Type: cross Abstract: Robots that operate over extended periods should not merely visit space; they should progressively understand it. Yet most 3D scene graph pipelines tr

agentsarxiv-cs-ai
8 Jun 2026
Agents

The AI champions strategy of 2023 doesn't work in 2026. Let me give you my very hot take 🔥 (And know that this is anecdotal, and the world …

DGX agent

The AI champions strategy of 2023 doesn't work in 2026. Let me give you my very hot take 🔥 (And know that this is anecdotal, and the world of AI changes every 2 heartbeats, so by the time I finish thi

agentsallie-k--miller--x
8 Jun 2026
Agents

AIを作るAIを作る:RSI Lab始動 https://sakana.ai/rsi-lab-jp/ Sakana AIは、再帰的自己改善(Recursive Self-Improvement、RSI)に取り組む専任の研究グループ「RSI Lab」を、東京で立ち上げます。RSIは…

DGX agent

AIを作るAIを作る:RSI Lab始動 https://sakana.ai/rsi-lab-jp/ Sakana AIは、再帰的自己改善(Recursive Self-Improvement、RSI)に取り組む専任の研究グループ「RSI Lab」を、東京で立ち上げます。RSIは、AIがAIそのものを作る仕組みです。 この2年間、私たちはLLM-Squared、Darwin Gödel Machi

agentsdavid-ha--x
7 Jun 2026
Agents

DAST: A VLM-LLM Framework for Cross-Interface Anomaly Detection in O-RAN

DGX agent

arXiv:2606.06261v1 Announce Type: cross Abstract: O-RAN enables a disaggregated baseband stack with programmable functions that communicate over standardized open interfaces. The same openness that en

agentsarxiv-cs-ai
6 Jun 2026
Agents

Hermes Desktop 现已支持简体中文——聊天界面完整适配简体中文。 桌面应用现在已在所有 UI 界面全面提供简体中文(Simplified Chinese / 简体中文)翻译:包括聊天窗口本身、侧边栏、设置、命令中心、cron、消息、个人资料、技能、智能体等全部内容。 …

DGX agent

Hermes Desktop 现已支持简体中文——聊天界面完整适配简体中文。 桌面应用现在已在所有 UI 界面全面提供简体中文(Simplified Chinese / 简体中文)翻译:包括聊天窗口本身、侧边栏、设置、命令中心、cron、消息、个人资料、技能、智能体等全部内容。 英语仍为默认语言;你可以在“外观”设置中切换语言,选择后会自动保存到配置项 display.language。 Herm

agentsnous-research--x
6 Jun 2026
Agents

This analysis from yesterday looks even more apt today — especially after we heard that the model providers have gone to Washington in searc…

DGX agent

This analysis from yesterday looks even more apt today — especially after we heard that the model providers have gone to Washington in search of handouts. ⚠️Why didn’t the hyperscalers wait until afte

agentsgary-marcus--x
6 Jun 2026
Agents

Executable Schema Contracts: From Automatic Ingestion to Multi-Source Retrieval

DGX agent

arXiv:2606.05415v1 Announce Type: new Abstract: Real-world data spans tables, documents, and semi-structured files with implicit semantics. Querying this data requires integrating evidence across inco

agentsarxiv-cs-cl
5 Jun 2026
Agents

The Download: AI hacking beyond Mythos, and chatbots’ impact on our brains

DGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. The Meta hack shows there’s more to AI security than Mythos On

agentsmit-tech-review
5 Jun 2026
Agents

1/ 🔥 @NoPriorsPod x @LatentSpacePod chat with @SatyaNadella at @Microsoft Build. He has the sharpest mental models of any public company CE…

DGX agent

1/ 🔥 @NoPriorsPod x @LatentSpacePod chat with @SatyaNadella at @Microsoft Build. He has the sharpest mental models of any public company CEO I've interviewed. $MSFT is at its heart still a tools compa

agentsswyx--x
4 Jun 2026
Model Releases

AutoLab: Can Frontier Models Solve Long-Horizon Auto Research and Engineering Tasks?

DGX agent

arXiv:2606.05080v1 Announce Type: new Abstract: Scientific and engineering progress is fundamentally a long-horizon iterative process: proposing changes, running experiments, measuring outcomes, and c

model-releasesarxiv-cs-ai
4 Jun 2026
Agents

ContactExplorer: Contact Coverage-Guided Exploration for General-Purpose Dexterous Manipulation

DGX agent

arXiv:2603.10971v2 Announce Type: replace-cross Abstract: Reinforcement learning has achieved remarkable success in domains such as Atari games, navigation, and locomotion, where exploration can often

agentsarxiv-cs-ai
4 Jun 2026
Model Releases

CoPark: Learning Reactive Parking via Self-Play

DGX agent

arXiv:2606.04149v1 Announce Type: new Abstract: Learning a single policy that reaches a goal with high geometric precision while interacting safely with nearby agents poses conflicting objectives. Pre

model-releasesarxiv-cs-ro
4 Jun 2026
Agents

Description-Code Inconsistency in Real-world MCP Servers: Measurement, Detection, and Security Implications

DGX agent

arXiv:2606.04769v1 Announce Type: cross Abstract: The Model Context Protocol (MCP) has emerged as a critical standard empowering Large Language Models (LLMs) to utilize external tools. In this ecosyst

agentsarxiv-cs-ai
4 Jun 2026
Agents

here’s why i think the situation may be dire: https://x.com/GaryMarcus/status/2062369417372258652?s=20

DGX agent

here’s why i think the situation may be dire: https://x.com/GaryMarcus/status/2062369417372258652?s=20 ⚠️Why didn’t the hyperscalers wait until after the IPO in switching to pay-by-usage-charging? My

agentsgary-marcus--x
4 Jun 2026
Agents

Many research labs only consider inference efficiency after the fact. Step 3.7 Flash is a 198B sparse MoE VLM designed by @StepFun_ai for in…

DGX agent

Many research labs only consider inference efficiency after the fact. Step 3.7 Flash is a 198B sparse MoE VLM designed by @StepFun_ai for inference from the start. 196B language backbone with a 1.8B v

agentsfireworks-ai--x
4 Jun 2026
Model Releases

NVIDIA’s Nemotron 3 Ultra is available on Ollama’s cloud! Try it 👇 Claude Code: ollama launch claude --model nemotron-3-ultra:cloud Hermes …

DGX agent

NVIDIA’s Nemotron 3 Ultra is available on Ollama’s cloud! Try it 👇 Claude Code: ollama launch claude --model nemotron-3-ultra:cloud Hermes Agent: ollama launch hermes --model nemotron-3-ultra:cloud Op

model-releasesollama--x
4 Jun 2026
Model Releases

Safety Under Scaffolding: How Evaluation Conditions Shape Measured Safety

DGX agent

arXiv:2603.10044v2 Announce Type: replace-cross Abstract: A safety score earned on a benchmark need not predict how the same model behaves once it is wrapped in an agentic scaffold the benchmark never

model-releasesarxiv-cs-ai
4 Jun 2026
Agents

SUSD: Structured Unsupervised Skill Discovery through State Factorization

DGX agent

arXiv:2602.01619v2 Announce Type: replace-cross Abstract: Unsupervised Skill Discovery (USD) aims to autonomously learn a diverse set of skills without relying on extrinsic rewards. One of the most co

agentsarxiv-cs-ai
4 Jun 2026
Agents

⚠️Why didn’t the hyperscalers wait until after the IPO in switching to pay-by-usage-charging? My guess is that *they literally could not aff…

DGX agent

⚠️Why didn’t the hyperscalers wait until after the IPO in switching to pay-by-usage-charging? My guess is that *they literally could not afford to* —because it would bankrupt them. My reasoning: in “a

agentsgary-marcus--x
4 Jun 2026
Agents

BotDirector: Robot Storytelling Across the Symmetrical Reality with Multi-modal Interactions

DGX agent

arXiv:2606.03223v1 Announce Type: cross Abstract: Robot storytelling offers a unique blend of technological innovation and creative expression that engages children in unprecedented ways. However, the

agentsarxiv-cs-ai
3 Jun 2026
Agents

Build smarter document workflows: What’s new in Azure Content Understanding at Build 2026

DGX agent

Azure Content Understanding (CU) in Foundry Tools is Microsoft’s comprehensive content AI service. It ingests diverse data types — documents, audio, images, and video — and extracts the most critical

agentsmicrosoft-foundry
3 Jun 2026
Model Releases

Demo2Tutorial: From Human Experience to Multimodal Software Tutorials

DGX agent

arXiv:2606.03951v1 Announce Type: new Abstract: Human experience in digital environments offers a vast, underexplored resource of authentic, untrimmed interactions that contain rich procedural knowled

model-releasesarxiv-cs-cv
3 Jun 2026
Agents

DyaPlex: Full-Duplex Speech-Motion Model for Dyadic Interaction

DGX agent

arXiv:2606.03874v1 Announce Type: new Abstract: We present DyaPlex, a streaming, full-duplex speech-and-motion model designed for dyadic interaction. To capture the continuous and reciprocal nature of

agentsarxiv-cs-cv
3 Jun 2026
Safety

Easy-to-Use Shielding for Reinforcement Learning

DGX agent

arXiv:2606.03804v1 Announce Type: new Abstract: Safe exploration is a key challenge in Reinforcement Learning (RL) that aims to prevent agents from making harmful decisions while exploring their envir

safetyarxiv-cs-lg
3 Jun 2026
Agents

Entropy Gate: Entropy Quenching for Near-Lossless Token Compression in LLM Pipelines

DGX agent

arXiv:2606.03739v1 Announce Type: new Abstract: LLM pipelines waste substantial token budgets on low-information content: repeated context, verbose responses, and redundant boilerplate. We introduce E

agentsarxiv-cs-cl
3 Jun 2026
Agents

From 'What' to 'How' and 'Why': Sharing LLM-Generated Retrospective Summaries of Older Adults' Passive Tracking Data with Remote Family Members

DGX agent

arXiv:2606.03876v1 Announce Type: cross Abstract: With the growing prevalence of modern ubiquitous computing technologies, multi-modal tracking systems hold promise for providing timely awareness and

agentsarxiv-cs-ai
3 Jun 2026
Agents

langsmith! ✅ Sandbox: https://docs.langchain.com/langsmith/sandboxes ✅ Gateway: https://docs.langchain.com/langsmith/llm-gateway ✅ Observabi…

DGX agent

langsmith! ✅ Sandbox: https://docs.langchain.com/langsmith/sandboxes ✅ Gateway: https://docs.langchain.com/langsmith/llm-gateway ✅ Observability: https://docs.langchain.com/langsmith/observability eve

agentsharrison-chase--x
3 Jun 2026
Safety

Post-Hoc Robustness for Model-Based Reinforcement Learning

DGX agent

arXiv:2606.03521v1 Announce Type: cross Abstract: To improve the real-world applicability of reinforcement learning (RL), the field of adversarially robust RL studies how to train agents under adversa

safetyarxiv-cs-ai
3 Jun 2026
Agents

Whom to Query for What: Adaptive Group Elicitation via Multi-Turn LLM Interactions

DGX agent

arXiv:2602.14279v2 Announce Type: replace-cross Abstract: Eliciting information to reduce uncertainty about latent group-level properties from surveys and other collective assessments requires allocat

agentsarxiv-cs-ai
3 Jun 2026
Agents

With LangSmith Engine, systemic issues get surfaced automatically instead of getting buried in traces. @ollieelmgren from @ListenLabs on how…

DGX agent

With LangSmith Engine, systemic issues get surfaced automatically instead of getting buried in traces. @ollieelmgren from @ListenLabs on how LangSmith Engine changed the way his team evaluates their a

agentsharrison-chase--x
3 Jun 2026
Agents

4D Radar Meets LiDAR and Camera: Cooperative Perception under Adverse Weather

DGX agent

arXiv:2606.00416v1 Announce Type: new Abstract: Cooperative perception is important for autonomous driving but remains fragile when cameras and LiDAR degrade in adverse weather. We address this challe

agentsarxiv-cs-cv
2 Jun 2026
Model Releases

AblationBench: Evaluating Automated Planning of Ablations in Empirical AI Research

DGX agent

arXiv:2507.08038v3 Announce Type: replace-cross Abstract: Language model agents are increasingly used to automate scientific research, yet evaluating their scientific contributions remains a challenge

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

AI-IoT-Robotics Integration: Survey of Frameworks, Emerging Trends, and the Path Toward Connected Robotics

DGX agent

arXiv:2606.01015v1 Announce Type: cross Abstract: The convergence of Artificial Intelligence, the Internet of Things, and Robotics is no longer a futuristic vision; it is rapidly becoming the foundati

agentsarxiv-cs-ai
2 Jun 2026
Agents

All Models are Wrong, Knowing Where is Useful: On Model Uncertainty in Reinforcement Learning

DGX agent

arXiv:2606.01363v1 Announce Type: new Abstract: Model-based reinforcement learning (MBRL) infers information about the environment from a learned dynamics model and bears the potential to address open

agentsarxiv-cs-lg
2 Jun 2026
Agents

Application of Algorithms in Energy-Efficient Design Platforms for Green Building

DGX agent

arXiv:2606.01229v1 Announce Type: new Abstract: During green building design, computer-aided energy assessment is widely used to improve efficiency and achieve overall optimization. This paper present

agentsarxiv-cs-ai
2 Jun 2026
Safety

Behavior-Invariant Task Representation Learning with Transformer-based World Models for Offline Meta-Reinforcement Learning

DGX agent

arXiv:2606.00780v1 Announce Type: cross Abstract: Offline meta-reinforcement learning leverages static datasets to enable agents to generalize to unseen environments by combining offline efficiency wi

safetyarxiv-cs-ai
2 Jun 2026
Agents

Come check out our booth at Snowflake Summit 2026 ❄️ We Parse PDFs

DGX agent

Come check out our booth at Snowflake Summit 2026 ❄️ We Parse PDFs ❄️ Come meet the LlamaIndex Team at Snowflake Summit 2026. It might be chilly in Snow Park ☃️ but the AI infrastructure market is red

agentsjerry-liu--x
2 Jun 2026
Agents

Comprehensive AI governance requires addressing non-model gains

DGX agent

arXiv:2606.00047v1 Announce Type: cross Abstract: Frontier AI governance often centres on the model-level governance paradigm, which assumes that a model's capability profile is primarily a function o

agentsarxiv-cs-ai
2 Jun 2026
Agents

CountGD++: Generalized Prompting for Open-World Counting

DGX agent

arXiv:2512.23351v2 Announce Type: replace Abstract: The flexibility and accuracy of methods for automatically counting objects in images and videos are limited by the way the object can be specified.

agentsarxiv-cs-cv
2 Jun 2026
Agents

Everywhere Learning: Artificial Intelligence with Pointwise Constraints

DGX agent

arXiv:2606.01557v1 Announce Type: new Abstract: Everywhere learning is a new paradigm whereby Artificial Intelligence (AI) systems are trained to satisfy loss constraints with probability one over the

agentsarxiv-cs-lg
2 Jun 2026
Model Releases

FVSpec: Real-World Property-Based Tests as Lean Challenges

DGX agent

arXiv:2606.01008v1 Announce Type: cross Abstract: We present a benchmark for evaluating AI models and agents on real-world formal software verification tasks. We first scrape 11,039 property-based tes

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

Generative AI and Digital Ecosystem Resilience: A Proactive Lifecycle-Based Survey

DGX agent

arXiv:2606.00136v1 Announce Type: cross Abstract: The proliferation of adversarial synthetic content, accelerated by Generative AI (GenAI) is rendering traditional reactive detection methods ineffecti

agentsarxiv-cs-ai
2 Jun 2026
Agents

Genotype-Conditioned Molecular Generation via Evidence-Grounded Multi-Objective Latent Perturbation in Diffusion Models

DGX agent

arXiv:2606.01461v1 Announce Type: new Abstract: Developing effective anticancer therapeutics remains challenging due to tumor heterogeneity and the absence of well-defined molecular targets across can

agentsarxiv-cs-lg
2 Jun 2026
← Previous
1…234235236237238…375
Next →