AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,958 results
1 Jun 2026

A Unified and Reproducible Experimentation Framework for Speech Understanding

AgentsDGX agent

arXiv:2605.30899v1 Announce Type: cross Abstract: Speech foundation models and Speech LLMs have advanced speech understanding, yet deployment-oriented model selection is hindered by non-comparable eva

❄️ Come meet the LlamaIndex Team at Snowflake Summit 2026. It might be chilly in Snow Park ☃️ but the AI infrastructure market is red hot 🔥…

AgentsDGX agent

❄️ Come meet the LlamaIndex Team at Snowflake Summit 2026. It might be chilly in Snow Park ☃️ but the AI infrastructure market is red hot 🔥. Come visit our team at the expo floor and explore how you c

Fighting Numerical Hallucinations via Data-centric Compilation for Online Financial QA

AgentsDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.31064v1 Announce Type: cross Abstract: Large Language Models (LLMs) have significantly advanced online data services, particularly in the domain of financial question answering (FinQA). How

longmemeval experiment arch: 1) deterministic ingestion/extraction (85.6% accuracy, 86.2% retrieval) 2) semantic ingestion/extraction (84.8%…

AgentsDGX agent

longmemeval experiment arch: 1) deterministic ingestion/extraction (85.6% accuracy, 86.2% retrieval) 2) semantic ingestion/extraction (84.8% accuracy, 94.9% retention) 3) semantic ingestion/determinis

MatchFixAgent: Language-Agnostic Autonomous Repository-Level Code Translation Validation and Repair

AgentsDGX agent

arXiv:2509.16187v3 Announce Type: replace-cross Abstract: Code translation transforms source code from one programming language (PL) to another. Validating the functional equivalence of translation an

May 2026 newsletter

AgentsDGX agent

I just sent out the May edition of my sponsors-only monthly newsletter. If you are a sponsor (or if you start a sponsorship now) you can access it here. This month: Al got expensive, and Anthropic had

MedCoG: Maximizing LLM Inference Density in Medical Reasoning via Meta-Cognitive Regulation

AgentsDGX agent

arXiv:2602.07905v2 Announce Type: replace Abstract: Large Language Models (LLMs) have shown strong potential in complex medical reasoning yet face diminishing gains under inference scaling laws. While

PictSure: Pretraining Embeddings Matters for In-Context Learning Image Classifiers

AgentsDGX agent

arXiv:2506.14842v2 Announce Type: replace-cross Abstract: Building image classification models remains cumbersome in data-scarce domains, where collecting large labeled datasets is impractical. In-con

Pull Requests as a Training Signal for Repo-Level Code Editing

AgentsDGX agent

arXiv:2602.07457v2 Announce Type: replace-cross Abstract: Repository-level code editing requires models to understand complex dependencies and execute precise multi-file modifications across a large c

Social Reasoning in Machines: Investigating Collective Truth-Seeking Dynamics in Large Language Model Debate

AgentsDGX agent

arXiv:2605.30391v1 Announce Type: cross Abstract: Human reasoning has long been theorised to operate socially, not through isolated individual cognition, but through collective adversarial discourse,

Subspace-Decomposed JEPAs: Disentangling Progression and Content in Latent World Models

AgentsDGX agent

arXiv:2605.31111v1 Announce Type: new Abstract: Joint-Embedding Predictive Architectures (JEPAs) learn compact latent world models by predicting future embeddings, but no single coordinate of the late

Survival Reinforcement Learning: Toward Scalable Self-Supervised RL

AgentsDGX agent

arXiv:2605.31273v1 Announce Type: new Abstract: While self-supervised Contrastive Reinforcement Learning (CRL) has shown remarkable depth-scaling capabilities, successfully using networks over 64 laye

This is great @hwchase17 @bryonkuchML Seeing this update, I’m building a tutorial repo around LangChain + Groq + GEPA. The idea is simple: •…

AgentsDGX agent

This is great @hwchase17 @bryonkuchML Seeing this update, I’m building a tutorial repo around LangChain + Groq + GEPA. The idea is simple: •LangChain builds the RAG/agent workflow •Groq gives fast inf

TUX: Measuring Human--AI Tacit Understanding

SafetyDGX agent

arXiv:2605.30930v1 Announce Type: cross Abstract: As large language models (LLMs) increasingly act as collaborative partners, human--AI alignment is often evaluated through explicit task success, accu

31 May 2026

love seeing tldraw's canvas sdk get absolutely maxed by replit team. Proud infra dad

AgentsDGX agent

love seeing tldraw's canvas sdk get absolutely maxed by replit team. Proud infra dad The best design work doesn't happen in a chat box. You need space to explore ideas, create variants, and iterate Me

The Top AI Papers of the Week (May 24 - May 31) - SkillOpt - AutoScientists - The Efficiency Frontier - Language Models Need Sleep - Adaptin…

AgentsDGX agent

The Top AI Papers of the Week (May 24 - May 31) - SkillOpt - AutoScientists - The Efficiency Frontier - Language Models Need Sleep - Adapting the Interface, Not the Model - Forecasting Scientific Prog

29 May 2026

A Deep Learning Model of Mental Rotation Informed by Interactive VR Experiments

AgentsDGX agent

arXiv:2512.13517v2 Announce Type: replace-cross Abstract: Mental rotation -- the ability to compare objects seen from different viewpoints -- is a fundamental example of mental simulation and spatial

DGSG-Mind: Dynamic 3D Gaussian Scene Graphs for Long-Term Scene Understanding and Grounding

AgentsDGX agent

arXiv:2605.29879v1 Announce Type: new Abstract: Integrating open-vocabulary semantic information into dynamic 3D scene representations is essential for long-term embodied scene understanding. However,

DOJ sues states that rejected ICE requests for undercover license plates

IndustryDGX agent

The Department of Justice filed lawsuits against Maine, Massachusetts, Oregon, and Washington state alleging their refusal to issue undercover license plates to federal agents imposes unconstitutional

From Prompts to Context: An Ontology-Driven Framework for Human-Generative AI Collaboration

AgentsDGX agent

arXiv:2605.29675v1 Announce Type: cross Abstract: Collaborations with Generative AI often begin with a short prompt and end with an opaque output, leaving implicit who was involved, what task was bein

Human-in-the-Loop Swarms: A Bionic Swarm Approach to Real-World Soil Mapping

AgentsDGX agent

arXiv:2605.29091v1 Announce Type: new Abstract: Swarm and field robotics face significant barriers to real-world validation due to the high cost and development time to deploy hardware. This paper int

MOOSE-Copilot: A Web-Based Interactive Assistant for Unified Exploratory and Fine-Grained Scientific Hypothesis Discovery

AgentsDGX agent

arXiv:2605.29475v1 Announce Type: cross Abstract: Large language models (LLMs) show remarkable potential in scientific hypothesis discovery. However, existing approaches face two critical limitations:

Offloading Score: Measuring AI Reliance Through Counterfactual Workflows

AgentsDGX agent

arXiv:2605.29392v1 Announce Type: cross Abstract: AI tools are increasingly integrated into real-world workflows. However, existing measures of reliance on these tools focus on AI output adoption or o

On Distributional Reinforcement Learning in Chaotic Dynamical Systems

AgentsDGX agent

arXiv:2605.30160v1 Announce Type: cross Abstract: Chaotic dynamical systems pose a fundamental challenge for Reinforcement Learning (RL): exponential sensitivity to initial conditions induces high-var

Opus 4.8 dropped today. ParseBench results are out. ✅ Slight gains: tables, semantic formatting, layout ⚠️ Slight regressions: charts, conte…

AgentsDGX agent

Opus 4.8 dropped today. ParseBench results are out. ✅ Slight gains: tables, semantic formatting, layout ⚠️ Slight regressions: charts, content faithfulness 💰 Slight price/page increase Lots of alpha l

Real-rootedness of the Poincare polynomials of overline{mathcal M}_{0,n}: an AI-assisted proof

AgentsDGX agent

arXiv:2605.29151v1 Announce Type: cross Abstract: We prove real-rootedness for the Poincare polynomial [ P_n(t)=sum_{i=0}^{n-3} im H^{2i}(overline{mathcal M}_{0,n};Q)t^i ] of the Deligne--Mumford modu

SchGen: PCB Schematic Generation with Semantic-Grounded Code Representations

AgentsDGX agent

arXiv:2605.30345v1 Announce Type: new Abstract: Printed circuit board (PCB) schematic design defines nearly all electronic hardware, but it remains manual and expertise-intensive. While generative AI

SEAL: Can Saturated Benchmarks Be Revived by LLM-as-a-Meta-Judge?

AgentsDGX agent

arXiv:2605.30104v1 Announce Type: new Abstract: Widely used language-model benchmarks are increasingly saturated, with frontier systems often receiving near-tied scores that standard metrics cannot re

This is a diary entry to myself, so I remember what AI was like today. It's just going to be a bullet-list stream of consciousness. - There …

Model ReleasesDGX agent

This is a diary entry to myself, so I remember what AI was like today. It's just going to be a bullet-list stream of consciousness. - There are still so many leaders that have never seen an agent run

VikingMem: A Memory Base Management System for Stateful LLM-based Applications

AgentsDGX agent

arXiv:2605.29640v1 Announce Type: new Abstract: Large Language Models have revolutionized interactive applications; however, their finite context windows pose a critical data management challenge for

VisualThink-VLA: Visual Intermediate Reasoning for Effective and Low-Latency Vision-Language-Action Policies

AgentsDGX agent

arXiv:2605.30011v1 Announce Type: cross Abstract: Recent work has begun to equip vision-language-action (VLA) policies with explicit intermediate reasoning. In embodied control, however, textual chain

28 May 2026

AI is changing everything — but Dell says data remains the enterprise’s crown jewel

AgentsDGX agent

The breakneck expansion of AI is forcing enterprise IT leaders to urgently rethink how infrastructure and cyber resilience work together to protect critical assets. Above all else, securing foundation

Chain-based Adaptive Reconfiguration Over Lattices for Hallucination Reduction

AgentsDGX agent

arXiv:2605.27706v1 Announce Type: new Abstract: We introduce CAROL (Chain-based Adaptive Reconfiguration Over Lattices), a probabilistic framework for test-time hallucination reduction in large langua

Detection Without Correction: A Two-Parameter Decomposition of Multi-Stage LLM Pipelines

Model ReleasesDGX agent

arXiv:2605.27559v1 Announce Type: cross Abstract: Multi-stage LLM pipelines that perform multi-agent debate, intrinsic self-correction, or retrieval-augmented verification exhibit puzzling aggregate b

Extrapolative Weight Averaging Reveals Correctness-Efficiency Frontiers in Code RL

AgentsDGX agent

arXiv:2605.28751v1 Announce Type: cross Abstract: Linear interpolation between fine-tuned checkpoints has been shown to trace the Pareto front between competing objectives, but whether extrapolative w

Large Language Models Approach Expert Pedagogical Quality in Math Tutoring but Differ in Instructional and Linguistic Profiles

AgentsDGX agent

arXiv:2512.20780v3 Announce Type: replace Abstract: Recent work has explored the use of large language models (LLMs) to generate tutoring responses in mathematics, yet it remains unclear how closely t

PIRS: Physics-Informed Reward Shaping for SAC-Based Building Energy Management

AgentsDGX agent

arXiv:2605.28232v1 Announce Type: new Abstract: Occupant comfort and grid-aware energy efficiency are competing objectives whose joint optimization depends critically on how reward functions are speci

SCALE-COMM: Shared, Contrastively-Aligned Latent Embeddings for MARL Communication

SafetyDGX agent

arXiv:2605.27532v1 Announce Type: new Abstract: Emergent communication enables partially observant Autonomous Mobile Robots (AMRs) to coordinate effectively in decentralized multi-agent reinforcement

Securing Retrieval-Augmented Generation: A Taxonomy of Attacks, Defenses, and Future Directions

AgentsDGX agent

arXiv:2604.08304v2 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) extends large language models (LLMs) with external knowledge, but this access path also introduces securi

The Energy Blind Spot: NVIDIA's Flagship Edge AI Hardware Cannot Support Process-Level Energy Attribution

Local AiDGX agent

arXiv:2605.27599v1 Announce Type: cross Abstract: Agentic AI workloads - where a single user goal triggers multi-step orchestration, tool calls, retries, and failure recovery - are being targeted for

Verifiable Benchmarking of Long-Horizon Spatial Biology

Model ReleasesDGX agent

arXiv:2605.28065v1 Announce Type: new Abstract: AI agents are increasingly useful for biological data analysis, but existing benchmarks mostly test broad biological knowledge, executable workflows, or

VibeSearchBench: Benchmarking Long-horizon Proactive Search in the Wild

Model ReleasesDGX agent

arXiv:2605.27882v1 Announce Type: cross Abstract: LLM-based agents score well on search benchmarks, yet real users consistently find results unsatisfying, revealing a persistent evaluation-experience

27 May 2026

Can Broad Biomedical Knowledge be Contextualized into Scenario-Grounded Propositions?

AgentsDGX agent

arXiv:2605.27082v1 Announce Type: new Abstract: Biomedical discovery often requires connecting broad biomedical knowledge with specific experimental or clinical data. Background knowledge suggests rel

Cost of Structural Learning Under Censored Feedback: A Threshold-Bandit Approach

SafetyDGX agent

arXiv:2605.27076v1 Announce Type: cross Abstract: In many multi-agent applications, tasks yield rewards only when executed by a coalition meeting an unknown size threshold; otherwise, feedback is full

Design First, Code Later: Aesthetically Pleasing Template-Free Slides Generation

AgentsDGX agent

arXiv:2605.26451v1 Announce Type: cross Abstract: Producing presentation slides automatically entails coordinating narrative structure with page-level graphic design under strict spatial constraints.

Disentangled Representation Learning through Unsupervised Symmetry Group Discovery

AgentsDGX agent

arXiv:2603.11790v3 Announce Type: replace Abstract: Symmetry-based disentangled representation learning leverages the group structure of environment transformations to uncover the latent factors of va

For future-proof, build AI that's composable. Regardless of what you use, all these should be composable, iterative, and customizable: - LLM…

AgentsDGX agent

For future-proof, build AI that's composable. Regardless of what you use, all these should be composable, iterative, and customizable: - LLMs - Evals - Automations - MCP/CLI tools - Skills/Memory/Cont

Learning GUI Grounding with Spatial Reasoning from Visual Feedback

AgentsDGX agent

arXiv:2509.21552v2 Announce Type: replace-cross Abstract: Graphical User Interface (GUI) grounding is commonly framed as a coordinate prediction task -- given a natural language instruction, generate

Neural Bayesian Sequential Routing

AgentsDGX agent

arXiv:2605.26147v1 Announce Type: new Abstract: Human decision-making is sequential and uncertainty-aware, yet standard neural networks often rely on static, dense forward computation with limited vis

Optimal Rates for Feasible Payoff Set Estimation in Games

AgentsDGX agent

arXiv:2602.04397v2 Announce Type: replace-cross Abstract: We study a setting in which two players play a (possibly approximate) Nash equilibrium of a bimatrix game, while a learner observes only their

Persona Generators: Generating Diverse Synthetic Personas for Arbitrary Contexts

AgentsDGX agent

arXiv:2602.03545v2 Announce Type: replace Abstract: Evaluating AI systems that interact with humans requires understanding their behavior across diverse user populations, but collecting representative

PolyFusionAgent: A Multimodal Foundation Model and Autonomous AI Assistant for Polymer Property Prediction and Inverse Design

AgentsDGX agent

arXiv:2605.26543v1 Announce Type: new Abstract: Polymer discovery is central to fields ranging from energy storage to biomedicine, but it is hindered by an astronomically large chemical design space a

Telenor Nordics Customer Service self-help corpus

AgentsDGX agent

arXiv:2605.26891v1 Announce Type: new Abstract: This paper presents a multilingual customer service self-help corpus comprising 1,122 manually validated documents in Finnish, Danish, Norwegian, and Sw

When Does Adaptive Guidance Help? Belief-Aware Privileged Distillation for Autonomous Driving Under Partial Observability

AgentsDGX agent

arXiv:2605.26155v1 Announce Type: cross Abstract: Guided Soft Actor-Critic (GSAC) distills knowledge from a privileged full-state teacher to a partial-observation student for autonomous driving, but u

WisdomAI brings plug-and-play conversational AI analytics capabilities to SaaS applications

AgentsDGX agent

Not content with its push into autonomous artificial intelligence workers, WisdomAI Inc. is bringing its comprehensive analytics capabilities to third-party software-as-a-service platforms with today’

@yoheinakajima is genuinely the coolest power user you can have as an ai company got to talk to him the other day about how he's using cofou…

ApplicationsDGX agent

@yoheinakajima is genuinely the coolest power user you can have as an ai company got to talk to him the other day about how he's using cofounder and left thinking about how i coddle my agents too much

26 May 2026

AGI Requires a Coordination Layer on Top of Pattern Repositories

AgentsDGX agent

arXiv:2512.05765v2 Announce Type: replace Abstract: In this paper we argue that influential critiques dismissing Large Language Models (LLMs) as a dead end for AGI misidentify the bottleneck: they con

Anisotropic Diffusion-Driven Ergodic Coverage in Multi-Robot Systems

AgentsDGX agent

arXiv:2605.24125v1 Announce Type: new Abstract: We consider the problem of combining potential field and ergodic search on multi-robot systems. Traditional ergodic search algorithms use metrics for er

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork

Model ReleasesDGX agent

arXiv:2605.24423v1 Announce Type: new Abstract: In-Context Reinforcement Learning (ICRL) has enabled foundation agents to adapt instantaneously to novel tasks, yet its efficacy in Ad-Hoc Teamwork (AHT

CVEvolve: Autonomous Algorithm Discovery for Unstructured Scientific Data Processing

AgentsDGX agent

arXiv:2605.11359v2 Announce Type: replace Abstract: Scientific data processing often requires task-specific algorithms or AI models, creating a barrier for domain scientists who need to analyze their

← Previous
1…188189190191192…300
Next →