AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,762 results
11 May 2026

MathlibPR: Pull Request Merge-Readiness Benchmark for Formal Mathematical Libraries

Model ReleasesDGX agent

arXiv:2605.07147v1 Announce Type: cross Abstract: The ecosystem of Lean and Mathlib has become the de facto standard for large language model (LLM) assisted formal reasoning with remarkable successes

Reason to Play: Behavioral and Brain Alignment Between Frontier LRMs and Human Game Learners

SafetyDGX agent

arXiv:2605.08019v1 Announce Type: new Abstract: Humans rapidly learn abstract knowledge when encountering novel environments and flexibly deploy this knowledge to guide efficient and intelligent actio

Scalable Option Learning in High-Throughput Environments

ResearchDGX agent

arXiv:2509.00338v3 Announce Type: replace-cross Abstract: Hierarchical reinforcement learning (RL) has the potential to enable effective decision-making over long timescales. Existing approaches, whil

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

SCENE: Recognizing Social Norms and Sanctioning in Group Chats

Model ReleasesDGX agent

arXiv:2605.07823v1 Announce Type: new Abstract: Online group chats are social spaces with implicit behavior patterns that, when broken, are often met with social sanctioning from the group. The abilit

Self Driving Datasets: From 20 Million Papers to Nuanced Biomedical Knowledge at Scale

Local AiDGX agent

arXiv:2605.07022v1 Announce Type: new Abstract: Manually curated biomedical repositories -- spanning bioactivity, genomics, and chemistry -- are expensive to maintain, lag behind primary literature, a

Towards Highly-Constrained Human Motion Generation with Retrieval-Guided Diffusion Noise Optimization

ResearchDGX agent

arXiv:2605.08054v1 Announce Type: new Abstract: Generating human motion that satisfies customized zero-shot goal functions, enabling applications such as controllable character animation and behavior

7 May 2026

Deco: Extending Personal Physical Objects into Pervasive AI Companion through a Dual-Embodiment Framework

ResearchDGX agent

arXiv:2605.03882v1 Announce Type: cross Abstract: Individuals frequently form deep attachments to physical objects (e.g., plush toys) that usually cannot sense or respond to their emotions. While AI c

Information Coordination as a Bridge: A Neuro-Symbolic Architecture for Reliable Autonomous Driving Scene Understanding

Model ReleasesDGX agent

arXiv:2605.04475v1 Announce Type: new Abstract: Reliable autonomous driving requires scene understanding that is semantically consistent across heterogeneous sensors and verifiable at the reasoning st

LUCAS-MEGA: A Large-Scale Multimodal Dataset for Representation Learning in Soil-Environment Systems

Model ReleasesDGX agent

arXiv:2605.04323v1 Announce Type: new Abstract: Understanding soil is fundamental to agriculture, carbon cycling, and environmental sustainability, yet progress is limited by fragmented and heterogene

MongoDB announces platform enhancements for enterprise-ready AI production

Local AiDGX agent

Popular NoSQL-based database company MongoDB Inc. today announced a new set of capabilities during the company’s .Local conference in London, bringing together everything software and artificial intel

Safety Must Precede the Deployment of Open-Ended AI

SafetyDGX agent

arXiv:2502.04512v3 Announce Type: replace Abstract: AI advancements have been significantly driven by a combination of foundation models and curiosity-driven learning aimed at increasing capability an

Saw this and thought 'yes! ChatGPT voice mode is going to stop acting like a two-year-model' but that upgrade hasn't shipped just yet

Model ReleasesDGX agent

Saw this and thought 'yes! ChatGPT voice mode is going to stop acting like a two-year-model' but that upgrade hasn't shipped just yet Introducing GPT-Realtime-2 in the API: our most intelligent voice

6 May 2026

A great conversation between Noah Kravitz from the @nvidia team + @hwchase17.

HardwareDGX agent

A great conversation between Noah Kravitz from the @nvidia team + @hwchase17. “Every enterprise needs a claw strategy.” How did @LangChain go from a weekend project to 1B+ downloads in 3 years? We sat

Learning Reactive Dexterous Grasping via Hierarchical Task-Space RL Planning and Joint-Space QP Control

SafetyDGX agent

arXiv:2605.03363v1 Announce Type: new Abstract: In this work, we propose a hybrid hierarchical control framework for reactive dexterous grasping that explicitly decouples high-level spatial intent fro

MCP-Atlas: A Large-Scale Benchmark for Tool-Use Competency with Real MCP Servers

Model ReleasesDGX agent

arXiv:2602.00933v2 Announce Type: replace-cross Abstract: The Model Context Protocol (MCP) is rapidly becoming the standard interface for Large Language Models (LLMs) to discover and invoke external t

MINT: Minimal Information Neuro-Symbolic Tree for Objective-Driven Knowledge-Gap Reasoning and Active Elicitation

SafetyDGX agent

arXiv:2602.05048v2 Announce Type: replace Abstract: Joint planning through language-based interactions is a key area of human-AI teaming. Planning problems in the open world often involve various aspe

Optimal control of the future via prospective learning with control

Model ReleasesDGX agent

arXiv:2511.08717v4 Announce Type: replace-cross Abstract: Optimal control of the future is the next frontier for AI. Current approaches to this problem are typically rooted in reinforcement learning (

Preemptive Solving of Future Problems: Multitask Preplay in Humans and Machines

TutorialsDGX agent

arXiv:2507.05561v2 Announce Type: replace Abstract: Humans can pursue a near-infinite variety of tasks, but typically can only pursue a small number at the same time. We hypothesize that humans levera

Yesterday @ElevenLabs announced they'd crossed >$500M ARR. Here's my interview with @matiii from @sequoia AI Ascent a few weeks ago. Mati is…

IndustryDGX agent

Yesterday @ElevenLabs announced they'd crossed >$500M ARR. Here's my interview with @matiii from @sequoia AI Ascent a few weeks ago. Mati is a founder who really inspires me. He shares stories about E

5 May 2026

AI Alignment via Incentives and Correction

SafetyDGX agent

arXiv:2605.01643v1 Announce Type: new Abstract: We study AI alignment through the lens of law-and-economics models of deterrence and enforcement. In these models, misconduct is not treated as an exter

For those affected, you're going to see a lot of advice on social media for what to do next. I like Claire's tweet a lot. Here's my version …

Model ReleasesDGX agent

For those affected, you're going to see a lot of advice on social media for what to do next. I like Claire's tweet a lot. Here's my version of her tweet for millennial/gen x business professionals: -

Microsoft's 2026 Work Trend Index: a survey of 20K AI users finds 65% fear falling behind, but only 13% report being rewarded for workplace AI experimentation (Todd Bishop/GeekWire)

IndustryDGX agent

Todd Bishop / GeekWire: Microsoft's 2026 Work Trend Index: a survey of 20K AI users finds 65% fear falling behind, but only 13% report being rewarded for workplace AI experimentation — [Editor's Note:

NaviMaster: Learning a Unified Policy for GUI and Embodied Navigation Tasks

SafetyDGX agent

arXiv:2508.02046v4 Announce Type: replace-cross Abstract: Recent advances in Graphical User Interface (GUI) and embodied navigation have driven progress, yet these domains have largely evolved in isol

The extit{Silicon Society} Cookbook: Design Space of LLM-based Social Simulations

HardwareDGX agent

arXiv:2605.00197v1 Announce Type: cross Abstract: Studies attempting to simulate human behavior with extit{Silicon Societies} grow in numbers while LLM-only social networks have started appearing outs

TRAP: Tail-aware Ranking Attack for World-Model Planning

SafetyDGX agent

arXiv:2605.01950v1 Announce Type: new Abstract: World models enable long-horizon planning by internally generating and evaluating imagined trajectories, making them a promising foundation for generali

4 May 2026

deepagents-cli is quietly becoming the best place to start coding with open weight models. we've been investing heavily in making it a harne…

Model ReleasesDGX agent

deepagents-cli is quietly becoming the best place to start coding with open weight models. we've been investing heavily in making it a harness that's truly model-agnostic, without compromising perform

Odysseus: Scaling VLMs to 100+ Turn Decision-Making in Games via Reinforcement Learning

ResearchDGX agent

arXiv:2605.00347v1 Announce Type: cross Abstract: Given the rapidly growing capabilities of vision-language models (VLMs), extending them to interactive decision-making tasks such as video games has e

Structure Liberates: How Constrained Sensemaking Produces More Novel Research Output

ResearchDGX agent

arXiv:2605.00557v1 Announce Type: new Abstract: Scientific discovery is an extended process of ideation--surveying prior work, forming hypotheses, and refining reasoning--yet existing approaches treat

3 May 2026

Wiki Lint Report — 2026-05-03

SynthesesDGX agent

Automated lint: 45 errors, 11 warnings, 3 info

1 May 2026

FineState-Bench: Benchmarking State-Conditioned Grounding for Fine-grained GUI State Setting

Model ReleasesDGX agent

arXiv:2604.27974v1 Announce Type: new Abstract: Despite the rapid progress of large vision-language models (LVLMs), fine-grained, state-conditioned GUI interaction remains challenging. Current evaluat

HAVEN: Hybrid Automated Verification ENgine for UVM Testbench Synthesis with LLMs

ResearchDGX agent

arXiv:2604.27643v1 Announce Type: cross Abstract: Integrated Circuit (IC) verification consumes nearly 70% of the IC development cycle, and recent research leverages Large Language Models (LLMs) to au

Intern-Atlas: A Methodological Evolution Graph as Research Infrastructure for AI Scientists

SafetyDGX agent

arXiv:2604.28158v1 Announce Type: new Abstract: Existing research infrastructure is fundamentally document-centric, providing citation links between papers but lacking explicit representations of meth

Learning-to-Explain through 20Q Gaming: An Explainable Recommender for Cybersecurity Education

SafetyDGX agent

arXiv:2604.26964v1 Announce Type: cross Abstract: The growing sophistication of contemporary cyber threats necessitates a more effective and adaptive approach to cybersecurity training. Intuitive and

Skills-Coach: A Self-Evolving Skill Optimizer via Training-Free GRPO

Model ReleasesDGX agent

arXiv:2604.27488v1 Announce Type: new Abstract: We introduce Skills-Coach, a novel automated framework designed to significantly enhance the self-evolution of skills within Large Language Model (LLM)-

SpatialGrammar: A Domain-Specific Language for LLM-Based 3D Indoor Scene Generation

Model ReleasesDGX agent

arXiv:2604.27555v1 Announce Type: new Abstract: Automatically generating interactive 3D indoor scenes from natural language is crucial for virtual reality, gaming, and embodied AI. However, existing L

30 Apr 2026

Omni2Sound: Towards Unified Video-Text-to-Audio Generation

Model ReleasesDGX agent

arXiv:2601.02731v3 Announce Type: replace-cross Abstract: Training a unified model integrating video-to-audio (V2A), text-to-audio (T2A), and joint video-text-to-audio (VT2A) generation offers signifi

Probe-then-Plan: Environment-Aware Planning for Industrial E-commerce Search

SafetyDGX agent

arXiv:2603.15262v2 Announce Type: replace Abstract: Modern e-commerce search is evolving to resolve complex user intents. While Large Language Models (LLMs) offer strong reasoning, existing LLM-based

The hidden risks of temporal resampling in clinical reinforcement learning

ApplicationsDGX agent

arXiv:2602.06603v3 Announce Type: replace Abstract: Reinforcement learning (RL) is a type of artificial intelligence for making optimal choices. In healthcare, researchers generally use offline RL (OR

29 Apr 2026

A Deep Reinforcement Learning Approach to Automated Stock Trading, using xLSTM Networks

SafetyDGX agent

arXiv:2503.09655v2 Announce Type: replace-cross Abstract: Traditional Long Short-Term Memory (LSTM) networks are effective for handling sequential data but have limitations such as gradient vanishing

A Survey on LLM-based Conversational User Simulation

ResearchDGX agent

arXiv:2604.24977v1 Announce Type: new Abstract: User simulation has long played a vital role in computer science due to its potential to support a wide range of applications. Language, as the primary

Certifyde raises $2M to help guide businesses in adopting and scaling AI

TutorialsDGX agent

Certifyde Inc. today announced it has raised 2 million to help businesses turn investment in artificial intelligence into broader adoption and fluency among rank and file. According to a 2025 report f

From World-Gen to Quest-Line: A Dependency-Driven Prompt Pipeline for Coherent RPG Generation

Local AiDGX agent

arXiv:2604.25482v1 Announce Type: new Abstract: Large Language Models (LLMs) have shown strong potential for narrative generation, but their use in complex, multi-layered role-playing game (RPG) world

Hightouch, which uses AI to help marketers create and manage campaigns, raised 150M led by Goldman Sachs and Bain at a 2.75B valuation, up from $1.2B in 2025 (Patrick Coffee/Wall Street Journal)

IndustryDGX agent

Patrick Coffee / Wall Street Journal: Hightouch, which uses AI to help marketers create and manage campaigns, raised 150M led by Goldman Sachs and Bain at a 2.75B valuation, up from 1.2B in 2025 — The

Learning from Medical Entity Trees: An Entity-Centric Medical Data Engineering Framework for MLLMs

SafetyDGX agent

arXiv:2604.25296v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have shown transformative potential in medical applications, yet their performance is hindered by conventional

OmniAlpha: Aligning Transparency-Aware Generation via Multi-Task Unified Reinforcement Learning

Local AiDGX agent

arXiv:2511.20211v2 Announce Type: replace Abstract: Transparency-aware generation requires modeling not only RGB appearance but also alpha-based opacity and cross-layer composition, which are essentia

VISION-SLS: Safe Perception-Based Control from Learned Visual Representations via System Level Synthesis

SafetyDGX agent

arXiv:2604.24894v1 Announce Type: cross Abstract: We propose VISION-SLS, a method for nonlinear output-feedback control from high-resolution RGB images which provides robust constraint satisfaction gu

Vocabulary Dropout for Curriculum Diversity in LLM Co-Evolution

SafetyDGX agent

arXiv:2604.03472v2 Announce Type: replace Abstract: Co-evolutionary self-play, where one language model generates problems and another solves them, promises autonomous curriculum learning without huma

28 Apr 2026

2026: the year of explosive AI innovation. @FireworksAI_HQ Co-founder & CEO Lin Qiao speaks on how the massive expansion of customized infer…

ApplicationsDGX agent

2026: the year of explosive AI innovation. @FireworksAI_HQ Co-founder & CEO Lin Qiao speaks on how the massive expansion of customized inference and models is creating speed of light production to sca

A Self-Supervised Framework for Space Object Behaviour Characterisation

SafetyDGX agent

arXiv:2504.06176v3 Announce Type: replace-cross Abstract: Foundation Models, which leverage large neural networks pre-trained on unlabelled data before fine-tuning for specific tasks, are increasingly

Aranya debuts cluster-scale operating system, partners with Hydra Host on ‘bare-metal AI’

HardwareDGX agent

Aranya Inc., a startup building a cluster-scale operating system built to meet demand for the next generation of supercomputer, launched today with major partnerships with top artificial intelligence

AWS brings OpenAI’s AI models and Codex programming assistant to its cloud

IndustryDGX agent

Amazon Web Services Inc. today made OpenAI Group PBC’s large language models available on its cloud platform. The algorithms are accessible through Amazon Bedrock alongside Codex, the ChatGPT develope

Bridging Reasoning and Action: Hybrid LLM-RL Framework for Efficient Cross-Domain Task-Oriented Dialogue

SafetyDGX agent

arXiv:2604.23345v1 Announce Type: new Abstract: Cross-domain task-oriented dialogue requires reasoning over implicit and explicit feasibility constraints while planning long-horizon, multi-turn action

CASP: Support-Aware Offline Policy Selection for Two-Stage Recommender Systems

SafetyDGX agent

arXiv:2604.23022v1 Announce Type: cross Abstract: Two-stage recommender systems first choose a candidate generator and then rank items within the generated set. Because the generator decides which ite

Context-Aware Hospitalization Forecasting Evaluations for Decision Support using LLMs

SafetyDGX agent

arXiv:2604.23949v1 Announce Type: new Abstract: Medical and public health experts must make real-time resource decisions, such as expanding hospital bed capacity, based on projected hospitalization tr

did a whole section on claude design because i'm loving it (like this deck) under the hood, it's claude code with an opinionated ontology on…

Model ReleasesDGX agent

did a whole section on claude design because i'm loving it (like this deck) under the hood, it's claude code with an opinionated ontology on both input (typography, logos, etc.) and output (slides, de

Do Synthetic Trajectories Reflect Real Reward Hacking? A Systematic Study on Monitoring In-the-Wild Hacking in Code Generation

SafetyDGX agent

arXiv:2604.23488v1 Announce Type: new Abstract: Reward hacking in code generation, where models exploit evaluation loopholes to obtain full reward without correctly solving the tasks, poses a critical

Efficient Rationale-based Retrieval: On-policy Distillation from Generative Rerankers based on JEPA

SafetyDGX agent

arXiv:2604.23336v1 Announce Type: cross Abstract: Unlike traditional fact-based retrieval, rationale-based retrieval typically necessitates cross-encoding of query-document pairs using large language

Explaining Temporal Graph Predictions With Shapley Values

Local AiDGX agent

arXiv:2604.24078v1 Announce Type: new Abstract: Temporal Graph Neural Networks (TGNNs) have become increasingly popular in recent years due to their superior predictive performance by combining both s

Fix Initial Codes and Iteratively Refine Textual Directions Toward Safe Multi-Turn Code Correction

SafetyDGX agent

arXiv:2604.23989v1 Announce Type: cross Abstract: Recent work on large language models (LLMs) has emphasized the importance of scaling inference compute. From this perspective, the state-of-the-art me

Fragmented AI policy threatens US leadership as government scrambles to keep pace

SafetyDGX agent

AI policy fragmentation is emerging as a critical risk for Washington, and without a federal standard, a patchwork of conflicting state-level rules threatens to undermine American competitiveness. Tha

← Previous
1…249250251252253…297
Next →