AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,607 results
Research

PhysForge: Generating Physics-Grounded 3D Assets for Interactive Virtual World

DGX agent

arXiv:2605.05163v1 Announce Type: new Abstract: Synthesizing physics-grounded 3D assets is a critical bottleneck for interactive virtual worlds and embodied AI. Existing methods predominantly focus on

researcharxiv-cs-cv
7 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

SCOUT: Active Information Foraging for Long-Text Understanding with Decoupled Epistemic States

DGX agent

arXiv:2605.04496v1 Announce Type: new Abstract: Long-Text Understanding (LTU) at million-token scale requires balancing reasoning fidelity with computational efficiency. Frontier long-context LLMs can

researcharxiv-cs-cl
7 May 2026
Tools

Searching for the right context in a sea of information is such a timeless problem in our field. I'm excited to talk about some of the appro…

DGX agent

Searching for the right context in a sea of information is such a timeless problem in our field. I'm excited to talk about some of the approaches we've learned and benefited from at Thrive next week a

toolslinus-lee--x
7 May 2026
Safety

To Fuse or to Drop? Dual-Path Learning for Resolving Modality Conflicts in Multimodal Emotion Recognition

DGX agent

arXiv:2605.04877v1 Announce Type: cross Abstract: Multimodal emotion recognition (MER) benefits from combining text, audio, and vision, yet standard fusion often fails when modalities conflict. Crucia

safetyarxiv-cs-lg
7 May 2026
Model Releases

We already had gemini-3.1-flash-lite-preview back on March 3rd, not clear if this new gemini-3.1-flash-lite is different other than no longe…

DGX agent

We already had gemini-3.1-flash-lite-preview back on March 3rd, not clear if this new gemini-3.1-flash-lite is different other than no longer being marked as a 'preview'. Pricing appears to be the sam

model-releasessimon-willison--x
7 May 2026
Model Releases

We've teamed up with @cerebras to offer free Windsurf plans for SWE-1.6 Fast Mode at up to 1000 tok/s! Fast Mode is built on Cerebras infere…

DGX agent

We've teamed up with @cerebras to offer free Windsurf plans for SWE-1.6 Fast Mode at up to 1000 tok/s! Fast Mode is built on Cerebras inference, enabling superior speed for planning and development wi

model-releaseswindsurf--x
7 May 2026
Model Releases

An explainable hypothesis-driven approach to Drug-Induced Liver Injury with HADES

DGX agent

arXiv:2605.02669v1 Announce Type: new Abstract: Drug-induced liver injury (DILI) remains a leading cause of late-stage clinical trial attrition. However, existing computational predictors primarily re

model-releasesarxiv-cs-ai
6 May 2026
Safety

Audio-Visual Intelligence in Large Foundation Models

DGX agent

arXiv:2605.04045v1 Announce Type: new Abstract: Audio-Visual Intelligence (AVI) has emerged as a central frontier in artificial intelligence, bridging auditory and visual modalities to enable machines

safetyarxiv-cs-cv
6 May 2026
Model Releases

Causal Software Engineering: A Vision and Roadmap

DGX agent

arXiv:2605.02454v1 Announce Type: cross Abstract: Software engineering increasingly involves making high-stakes decisions under uncertainty, using signals from code, field data, and socio-technical pr

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

DiagramNet: An End-to-End Recognition Framework and Dataset for Non-Standard System-Level Diagrams

DGX agent

arXiv:2605.01338v1 Announce Type: new Abstract: System-level diagrams encode the architectural blueprint of chip design, specifying module functions, dataflows, and interface protocols. However, non-s

model-releasesarxiv-cs-ai
6 May 2026
Safety

False Friends in the Shell: Unveiling the Emoticon Semantic Confusion in Large Language Models

DGX agent

arXiv:2601.07885v2 Announce Type: replace-cross Abstract: Emoticons are widely used in digital communication to convey affective intent, yet their safety implications for Large Language Models (LLMs)

safetyarxiv-cs-ai
6 May 2026
Model Releases

HAAS: A Policy-Aware Framework for Adaptive Task Allocation Between Humans and Artificial Intelligence Systems

DGX agent

arXiv:2605.02832v1 Announce Type: new Abstract: Deciding how to distribute work between humans and AI systems is a central challenge in organisational design. Most approaches treat this as a binary ch

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

Live blog: Code w/ Claude 2026

DGX agent

Simon Willison's live blog covers Anthropic's Code w/ Claude 2026 event, documenting the morning keynote sessions with real-time updates. The event featured announcements including updates to Claude m

model-releasessimon-willison
6 May 2026
Safety

Model Spec Midtraining: Improving How Alignment Training Generalizes

DGX agent

arXiv:2605.02087v1 Announce Type: new Abstract: Some frontier AI developers aim to align language models to a Model Spec or Constitution that describes the intended model behavior. However, standard a

safetyarxiv-cs-ai
6 May 2026
Model Releases

On Verbalized Confidence Scores for LLMs

DGX agent

arXiv:2412.14737v2 Announce Type: replace Abstract: The rise of large language models (LLMs) and their tight integration into our daily life make it essential to dedicate efforts towards their trustwo

model-releasesarxiv-cs-cl
6 May 2026
Industry

OpenAI phone leaks: a push toward AI first hardware.

DGX agent

OpenAI is fast-tracking development of its first AI-focused smartphone with mass production targeted for early 2027 , marking the company's entry into consumer hardware . The device, positioned as an

industryr-chatgpt
6 May 2026
Model Releases

PIIGuard: Mitigating PII Harvesting under Adversarial Sanitization

DGX agent

arXiv:2605.03129v1 Announce Type: cross Abstract: Browsing-enabled LLM assistants can fetch webpages and answer contact-seeking queries, creating a practical channel for scraping contact-style persona

model-releasesarxiv-cs-cl
6 May 2026
Applications

@PPLXfinance @PPLXDevs On FinSearchComp T1, Finance Search delivered the highest accuracy for live financial data and the lowest cost per co…

DGX agent

@PPLXfinance @PPLXDevs On FinSearchComp T1, Finance Search delivered the highest accuracy for live financial data and the lowest cost per correct answer in the cohort. Every result includes citations,

applicationsperplexity--x
6 May 2026
Model Releases

this deepagents deploy https://docs.langchain.com/oss/python/deepagents/deploy (or at least directionally where we want to take it) what's m…

DGX agent

this deepagents deploy https://docs.langchain.com/oss/python/deepagents/deploy (or at least directionally where we want to take it) what's missing? give us feedback! can someone PLEASE launch OS claud

model-releasesharrison-chase--x
6 May 2026
Model Releases

Two weeks after release, Hy3 preview is #1 on @OpenRouter's weekly leaderboard with 3.66T tokens processed, up 298% week-over-week. #1 in ov…

DGX agent

Two weeks after release, Hy3 preview is #1 on @OpenRouter's weekly leaderboard with 3.66T tokens processed, up 298% week-over-week. #1 in overall usage, tool calls, and coding. 15.4% market share acro

model-releasesjeremy-howard--x
6 May 2026
Model Releases

Valley3: Scaling Omni Foundation Models for E-commerce

DGX agent

arXiv:2605.01278v1 Announce Type: new Abstract: In this work, we present Valley3, an omni multimodal large language model (MLLM) developed for diverse global e-commerce tasks, with unified understandi

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

Vibe Code Bench: Evaluating AI Models on End-to-End Web Application Development

DGX agent

arXiv:2603.04601v2 Announce Type: replace-cross Abstract: Code generation has emerged as one of AI's highest-impact use cases, yet existing benchmarks measure isolated tasks rather than the complete '

model-releasesarxiv-cs-cl
6 May 2026
Local Ai

We built a simple app that's also probably the best PDF -> text phone app there. Take a picture of any document: a filled out form, identifi…

DGX agent

We built a simple app that's also probably the best PDF -> text phone app there. Take a picture of any document: a filled out form, identification, a statement, an essay - and we'll convert it into we

local-aijerry-liu--x
6 May 2026
Model Releases

we're continuing to see clear examples where a model's harness is a major determinant of overall performance. with the same model, running o…

DGX agent

we're continuing to see clear examples where a model's harness is a major determinant of overall performance. with the same model, running on same task, it's easy to observe very different scores depe

model-releasesharrison-chase--x
6 May 2026
Model Releases

Accelerating battery research with an AI interface between FINALES and Kadi4Mat

DGX agent

arXiv:2605.00909v1 Announce Type: cross Abstract: The time-consuming formation process critically impacts the longevity of sodium-ion coin cells and End Of Life (EOL) performance. This study aims to o

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Accurate Legal Reasoning at Scale: Neuro-Symbolic Offloading and Structural Auditability for Robust Legal Adjudication

DGX agent

arXiv:2605.02472v1 Announce Type: new Abstract: Legal texts often contain computational legal clauses--provisions whose understanding requires complex logic. While frontier Large Reasoning Models (LRM

model-releasesarxiv-cs-cl
5 May 2026
Research

Active Reasoning Vision-Language Models via Sequential Experimental Design

DGX agent

arXiv:2605.01345v1 Announce Type: new Abstract: Visual perception in modern Vision-Language Models (VLMs) is constrained by a fundamental perceptual bandwidth bottleneck: a broad field of view inevita

researcharxiv-cs-cv
5 May 2026
Safety

ARGUS: Policy-Adaptive Ad Governance via Evolving Reinforcement with Adversarial Umpiring

DGX agent

arXiv:2605.02200v1 Announce Type: new Abstract: Online advertising governance faces significant challenges due to the non-stationary nature of regulatory policies, where emerging mandates (e.g., restr

safetyarxiv-cs-cl
5 May 2026
Safety

Artificial intelligence language technologies in multilingual healthcare: Grand challenges ahead

DGX agent

arXiv:2605.01441v1 Announce Type: new Abstract: AI language technologies (AILTs), increasingly enabled by large language models (LLMs), are becoming embedded in multilingual healthcare workflows for t

safetyarxiv-cs-cl
5 May 2026
Model Releases

BIM Information Extraction Through LLM-based Adaptive Exploration

DGX agent

arXiv:2605.01698v1 Announce Type: new Abstract: BIM models provide structured representations of building geometry, semantics, and topology, yet extracting specific information from them remains remar

model-releasesarxiv-cs-cl
5 May 2026
Safety

Combining Trained Models in Reinforcement Learning

DGX agent

arXiv:2605.02159v1 Announce Type: new Abstract: Deep reinforcement learning (DRL) has delivered strong results in domains such as Atari and Go, but it still suffers from high sample cost and weak tran

safetyarxiv-cs-lg
5 May 2026
Model Releases

Decompose and Recompose: Reasoning New Skills from Existing Abilities for Cross-Task Robotic Manipulation

DGX agent

arXiv:2605.01448v1 Announce Type: cross Abstract: Cross-task generalization is a core challenge in open-world robotic manipulation, and the key lies in extracting transferable manipulation knowledge f

model-releasesarxiv-cs-cv
5 May 2026
Safety

DeepStage: Learning Autonomous Defense Policies Against Multi-Stage APT Campaigns

DGX agent

arXiv:2603.16969v2 Announce Type: replace-cross Abstract: This paper presents DeepStage, a deep reinforcement learning (DRL) framework for adaptive and stage-aware defense against Advanced Persistent

safetyarxiv-cs-lg
5 May 2026
Local Ai

DynoSLAM: Dynamic SLAM with Generative Graph Neural Networks for Real-World Social Navigation

DGX agent

arXiv:2605.02759v1 Announce Type: cross Abstract: Traditional Simultaneous Localization and Mapping (SLAM) algorithms rely heavily on the static environment assumption, which severely limits their app

local-aiarxiv-cs-cv
5 May 2026
Safety

Experience Constrained Hierarchical Federated Reinforcement Learning for Large-scale UAV Teams in Hazardous Environments

DGX agent

arXiv:2605.02165v1 Announce Type: new Abstract: Conventional federated learning assumes that greater learner participation improves training performance, by leveraging abundant, independently generate

safetyarxiv-cs-lg
5 May 2026
Model Releases

G-reasoner: Foundation Models for Unified Reasoning over Graph-structured Knowledge

DGX agent

arXiv:2509.24276v4 Announce Type: replace Abstract: Large language models (LLMs) excel at complex reasoning but remain limited by static and incomplete parametric knowledge. Retrieval-augmented genera

model-releasesarxiv-cs-ai
5 May 2026
Safety

Green Energy Management for Sustainable Data Centers Using Deep Reinforcement Learning

DGX agent

arXiv:2507.21153v2 Announce Type: replace Abstract: The exponential growth of digital services has positioned data centers among the most energy-intensive infrastructures in the modern economy, raisin

safetyarxiv-cs-lg
5 May 2026
Safety

High entropy leads to symmetry equivariant policies in Dec-POMDPs

DGX agent

arXiv:2511.22581v3 Announce Type: replace Abstract: We prove that in any Dec-POMDP, sufficiently high entropy regularization ensures that the policy gradient flow with tabular softmax parametrization

safetyarxiv-cs-lg
5 May 2026
Safety

Hybrid Quantum Reinforcement Learning with QAOA for Improved Vehicle Routing Optimization

DGX agent

arXiv:2605.01574v1 Announce Type: new Abstract: Vehicle Routing Problem (VRP) is one of the most complex NP-hard combinatorial optimization problem in transportation and logistics that requires a dyna

safetyarxiv-cs-lg
5 May 2026
Model Releases

i have yet to meet a single person who feels like claude code is getting exponentially better on some kind of fast take off

DGX agent

i have yet to meet a single person who feels like claude code is getting exponentially better on some kind of fast take off Anthropic pays $750K/ year per senior engineer. The creator of Claude Code j

model-releasesjeremy-howard--x
5 May 2026
Safety

Knowledge-Based Design Requirements for Generative Social Robots in Higher Education

DGX agent

arXiv:2602.12873v4 Announce Type: replace-cross Abstract: Generative social robots (GSRs) powered by large language models enable adaptive, conversational tutoring but also introduce risks such as mis

safetyarxiv-cs-ai
5 May 2026
Model Releases

MolViBench: Evaluating LLMs on Molecular Vibe Coding

DGX agent

arXiv:2605.02351v1 Announce Type: new Abstract: Molecular Vibe Coding, a paradigm where chemists interact with LLMs to generate executable programs for molecular tasks, has emerged as a flexible alter

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

OpenAI GPT-5 System Card

DGX agent

arXiv:2601.03267v2 Announce Type: replace Abstract: This is the system card published alongside the OpenAI GPT-5 launch, August 2025. GPT-5 is a unified system with a smart and fast model that answers

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Orchestrating Spatial Semantics via a Zone-Graph Paradigm for Intricate Indoor Scene Generation

DGX agent

arXiv:2605.02537v1 Announce Type: new Abstract: Autonomous 3D indoor scene synthesis breaks down in non-convex rooms with tightly coupled spatial constraints. Data-driven generators lack topological p

model-releasesarxiv-cs-ro
5 May 2026
Applications

Our AI started a cafe in Stockholm

DGX agent

Our AI started a cafe in Stockholm Andon Labs previously started an AI-run retail store in San Francisco. Now they're running a similar experiment in Stockholm, Sweden, only this time it's a cafe. The

applicationssimon-willison
5 May 2026
Model Releases

PACE: Parameter Change for Unsupervised Environment Design

DGX agent

arXiv:2605.01358v1 Announce Type: new Abstract: Unsupervised Environment Design (UED) offers a promising paradigm for improving reinforcement learning generalization by adaptively shaping training env

model-releasesarxiv-cs-lg
5 May 2026
Applications

RAST-MoE-RL: A Regime-Aware Spatio-Temporal MoE Framework for Deep Reinforcement Learning in Ride-Hailing

DGX agent

arXiv:2512.13727v2 Announce Type: replace Abstract: Ride-hailing platforms face the challenge of balancing passenger waiting times with overall system efficiency under highly uncertain supply-demand c

applicationsarxiv-cs-lg
5 May 2026
Safety

Reliability-Oriented Multilingual Orthopedic Diagnosis: A Domain-Adaptive Modeling and a Conceptual Validation Framework

DGX agent

arXiv:2605.02266v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly proposed for clinical decision support including multilingual diagnosis in low-resource settings. However,

safetyarxiv-cs-cl
5 May 2026
← Previous
1…351352353354355…367
Next →