AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,770 results
Model Releases

CADET: A Modular Platform for Evaluating Distributed Cooperative Autonomy in Connected Autonomous Vehicles

DGX agent

arXiv:2606.04072v1 Announce Type: cross Abstract: Deep learning models are increasingly central to autonomous vehicle (AV) pipelines, yet their integration has traditionally followed a monolithic desi

model-releasesarxiv-cs-lg
4 Jun 2026
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Caliper: Probing Lexical Anchors versus Causal Structure in LLMs

DGX agent

arXiv:2606.04915v1 Announce Type: new Abstract: Large language models reach 50 to 70% accuracy on causal reasoning benchmarks such as CLadder, but it is unclear whether this reflects structural reason

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

Can Generalist Agents Automate Data Curation?

DGX agent

arXiv:2606.04261v1 Announce Type: new Abstract: Curating training data is among the most consequential yet labor-intensive parts of modern AI development: practitioners iteratively propose, implement,

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Can I Take Another Dose? Evaluating LLM Decision-Making Under Temporal Uncertainty in OTC Dosing QA

DGX agent

arXiv:2606.04262v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used for everyday health questions, including whether a user can safely take another dose of an over-the

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Can Large Language Models Generalize Procedures Across Representations?

DGX agent

arXiv:2602.03542v2 Announce Type: replace Abstract: Large language models (LLMs) are trained and tested extensively on symbolic representations such as code and graphs, yet real-world user tasks are o

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

Caught in the Act(ivation): Toward Pre-Output and Multi-Turn Detection of Credential Exfiltration by LLM Agents

DGX agent

arXiv:2606.04141v1 Announce Type: cross Abstract: LLM agents often place sensitive credentials in the same context window as untrusted retrieved content, creating a direct path for indirect prompt inj

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

CDPM-Align: Multi-Scale Guidance-Aligned Diffusion Pretraining for Robust Few-Shot Anatomical Landmark Detection

DGX agent

arXiv:2606.04898v1 Announce Type: new Abstract: Anatomical landmark detection is a fundamental task in medical image analysis supporting a wide range of diagnostic and interventional workflows. Althou

model-releasesarxiv-cs-cv
4 Jun 2026
Model Releases

ChannelTok: Efficient Flexible-Length Vision Tokenization

DGX agent

arXiv:2606.04461v1 Announce Type: new Abstract: Leading flexible vision tokenizers achieve SOTA quality at an extreme cost, relying on parameter-heavy backbones and slow, multi-step generative decoder

model-releasesarxiv-cs-cv
4 Jun 2026
Model Releases

ChessMimic: Per-Rating Transformer Models for Human Move, Clock, and Outcome Prediction in Online Blitz Chess

DGX agent

arXiv:2606.04473v1 Announce Type: cross Abstract: We present ChessMimic, a system of three small encoder-only transformers - for move, thinking-time, and outcome prediction - conditioned on the positi

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

CodegenBench: Can LLMs Write Efficient Code Across Architectures?

DGX agent

arXiv:2606.04023v1 Announce Type: cross Abstract: While large language models (LLMs) have been extensively evaluated on code generation tasks for general-purpose programming and GPU-accelerated enviro

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

COMBINER: Composed Image Retrieval Guided by Attribute-based Neighbor Relations

DGX agent

arXiv:2606.04604v1 Announce Type: new Abstract: Composed Image Retrieval (CIR) represents a challenging retrieval task that targets locating specific images through multimodal inputs. Despite recent p

model-releasesarxiv-cs-cv
4 Jun 2026
Model Releases

Constraint-Enhanced Physical Search through Correlation Matching

DGX agent

arXiv:2606.03554v1 Announce Type: cross Abstract: Physical systems do not merely add noise to search processes; they impose constraints that generate structured correlations. We propose a principle of

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Continual Visual and Verbal Learning Through a Child's Egocentric Input

DGX agent

arXiv:2606.05115v1 Announce Type: cross Abstract: Children learn the meanings of words from a continuous, temporally structured stream of egocentric experience. Recent work shows that neural networks

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

CoPark: Learning Reactive Parking via Self-Play

DGX agent

arXiv:2606.04149v1 Announce Type: new Abstract: Learning a single policy that reaches a goal with high geometric precision while interacting safely with nearby agents poses conflicting objectives. Pre

model-releasesarxiv-cs-ro
4 Jun 2026
Model Releases

Curvature-aware dynamic precision approach for physics-informed neural networks

DGX agent

arXiv:2606.04736v1 Announce Type: cross Abstract: Physics-informed neural networks (PINNs) have become a promising framework for simulating partial differential equations (PDEs) by embedding physical

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

CyberGym-E2E: Scalable Real-World Benchmark for AI Agents' End-to-End Cybersecurity Capabilities

DGX agent

arXiv:2606.04460v1 Announce Type: cross Abstract: AI has the potential to transform cybersecurity by enabling systems that can autonomously detect, analyze, and remediate software vulnerabilities. How

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

D^3-MoE:Dual Disentangled Diffusion Mixture-of-Experts for Style-Controllable End-to-End Autonomous Driving

DGX agent

arXiv:2606.04884v1 Announce Type: new Abstract: Traditional end-to-end autonomous driving frameworks frequently suffer from the 'style-averaging' dilemma when trained on high-variance human demonstrat

model-releasesarxiv-cs-ro
4 Jun 2026
Model Releases

DeepSeek is becoming more popular among US enterprises as companies look for cheaper alternatives to Anthropic and OpenAI “DeepSeek takes to…

DGX agent

DeepSeek is becoming more popular among US enterprises as companies look for cheaper alternatives to Anthropic and OpenAI “DeepSeek takes top spot on 'trending' list as companies look for alternatives

model-releasesclem-delangue--x
4 Jun 2026
Model Releases

DetectZoo: A Unified Toolkit for AI-Generated Content Detection Across Text, Audio, and Image Modalities

DGX agent

arXiv:2606.04205v1 Announce Type: cross Abstract: The growing popularity and capacity of generative models have eroded the distinction between human and machine-generated content, motivating a growing

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Discourse-Role Labels as Presentation-Time Variables for Context Use in Language Models

DGX agent

arXiv:2606.04109v1 Announce Type: new Abstract: Context-augmented language model systems often wrap supplied content with labels such as Reference:, Evidence:, Instruction:, Note:, or Example:, but th

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

Dive into the Scene: Breaking the Perceptual Bottleneck in Vision-Language Decision Making via Focus Plan Generation

DGX agent

arXiv:2606.04046v1 Announce Type: cross Abstract: In embodied vision-language decision making tasks such as robotic manipulation and navigation, Vision-Language and Vision-Language-Action Models (VLMs

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

DLLG: Dynamic Logit-Level Gating of LLM Experts

DGX agent

arXiv:2606.04378v1 Announce Type: new Abstract: Leveraging multiple specialized LLMs can combine complementary strengths, but existing approaches trade adaptability for stability: routing commits prem

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

DLO-Lab: Benchmarking Deformable Linear Object Manipulations with Differentiable Physics

DGX agent

arXiv:2606.04206v1 Announce Type: new Abstract: We address the challenge of enabling robots to manipulate deformable linear objects (DLOs), such as ropes, cables, and rubber bands. Prior work has prim

model-releasesarxiv-cs-ro
4 Jun 2026
Model Releases

DMAConv: Dual Mask-Adaptive Convolution for Remote Sensing Pansharpening

DGX agent

arXiv:2512.08331v2 Announce Type: replace Abstract: Pansharpening aims to fuse a high-resolution panchromatic image with a low-resolution multispectral image. Existing deep learning methods, including

model-releasesarxiv-cs-cv
4 Jun 2026
Model Releases

dMX: Differentiable Mixed-Precision Assignment for Low-Precision Floating-Point Formats

DGX agent

arXiv:2606.04115v1 Announce Type: cross Abstract: Quantizing large language models (LLMs) to low-precision floating-point representations is central to efficient deployment, yet applying a single bit-

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Do Transformers Need Three Projections? Systematic Study of QKV Variants

DGX agent

arXiv:2606.04032v1 Announce Type: cross Abstract: Transformers have become the standard solution for various AI tasks, with the query, key, and value (QKV) attention formulation playing a central role

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Dream.exe: Can Video Generation Models Dream Executable Robot Manipulation?

DGX agent

arXiv:2606.04811v1 Announce Type: new Abstract: Video generation models have made impressive strides in synthesizing visually compelling content, yet their outputs remain confined to the virtual domai

model-releasesarxiv-cs-cv
4 Jun 2026
Model Releases

Dreaming: Better memory for a more helpful ChatGPT

DGX agent

OpenAI's 'Dreaming' feature enhances ChatGPT's memory capabilities by allowing the model to retain and leverage information from previous conversations to provide more personalized and contextually aw

model-releasesopenai
4 Jun 2026
Model Releases

Ekka: Automated Diagnosis of Silent Errors in LLM Inference

DGX agent

arXiv:2606.04594v1 Announce Type: cross Abstract: LLM serving frameworks are quickly evolving with a complex software stack and a vast number of optimizations. The rapid development process can introd

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Emotion Entanglement and Bayesian Inference for Multi-Dimensional Emotion Understanding

DGX agent

arXiv:2604.00819v2 Announce Type: replace-cross Abstract: Understanding emotions in natural language is inherently a multi-dimensional reasoning problem, where multiple affective signals interact thro

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Evaluating Large Language Models in Dynamic Clinical Decision-Making with Standardized Patient Cases

DGX agent

arXiv:2606.05112v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly proposed as clinical agents, yet static, single-turn benchmarks cannot capture how a model dynamically del

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

Evaluating Zero-Shot and One-Shot Adaptation of Small Language Models in Leader-Follower Interaction

DGX agent

arXiv:2602.23312v3 Announce Type: replace-cross Abstract: Leader-follower interaction is an important paradigm in human-robot interaction (HRI). Yet, assigning roles in real time remains challenging f

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

EvoPrompt: Guided Prompt Evolution for Vision-Language Models Adaptation

DGX agent

arXiv:2603.09493v2 Announce Type: replace-cross Abstract: The adaptation of large-scale vision-language models (VLMs) to downstream tasks with limited labeled data remains a significant challenge. Whi

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Exploring the Topology and Memory of Consensus: How LLM Agents Agree, Fragment, or Settle When Forming Conventions

DGX agent

arXiv:2606.04197v1 Announce Type: cross Abstract: How much should an LLM agent remember, and how should multi-agent systems be connected when trying to reach consensus? We show these two design choice

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

Exposing Blindspots: Cultural Bias Evaluation in Generative Image Models

DGX agent

arXiv:2510.20042v3 Announce Type: replace Abstract: Generative image models produce striking visuals yet often misrepresent culture. Prior work has examined cultural bias mainly in text-to-image (T2I)

model-releasesarxiv-cs-cv
4 Jun 2026
Model Releases

Finally! the first eval ship from cog!!!!!!!!!! 👼🏼 To contextualize: @METR_Evals cap out at ~16 hours. Cog has private enterprise evals up…

DGX agent

Finally! the first eval ship from cog!!!!!!!!!! 👼🏼 To contextualize: @METR_Evals cap out at ~16 hours. Cog has private enterprise evals up to 100hrs, and is confident enough to put a financial guarant

model-releasesswyx--x
4 Jun 2026
Model Releases

FindIt: A Format-Informed Visual Detection Benchmark for Generalist Multimodal LLMs

DGX agent

arXiv:2606.04282v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) are predominantly evaluated on free-form vision-language tasks such as visual question answering, captioning, a

model-releasesarxiv-cs-cv
4 Jun 2026
Model Releases

FinTradeBench: A Financial Reasoning Benchmark for LLMs

DGX agent

arXiv:2603.19225v3 Announce Type: replace-cross Abstract: Real-world financial decision-making is a challenging problem that requires reasoning over heterogeneous signals, including company fundamenta

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Flow Matching Calibration for Simulation-Based Inference under Model Misspecification

DGX agent

arXiv:2509.23385v5 Announce Type: replace-cross Abstract: Simulation-based inference (SBI) is transforming experimental sciences by enabling parameter estimation in complex non-linear models from simu

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

Food-R1: A Unified Multi-Task Food Vision-Language Model with Reinforcement Learning

DGX agent

arXiv:2606.04986v1 Announce Type: new Abstract: Recent studies have explored Vision-Language Models (VLMs) for food analysis. However, most existing methods rely primarily on supervised fine-tuning (S

model-releasesarxiv-cs-cv
4 Jun 2026
Model Releases

foom!

DGX agent

foom! Our internal data shows Claude is accelerating AI development—a possible path to recursive self-improvement, or AI autonomously building a more capable successor. It’s happening faster than we t

model-releasesemad-mostaque--x
4 Jun 2026
Model Releases

Founders Fund launches a TV-style game show featuring A-list founders and investors, including Sam Altman and Palmer Luckey, playing a game of Mafia (Tom Dotan/Newcomer)

DGX agent

Tom Dotan / Newcomer: Founders Fund launches a TV-style game show featuring A-list founders and investors, including Sam Altman and Palmer Luckey, playing a game of Mafia — Do people want to watch the

model-releasestechmeme
4 Jun 2026
Model Releases

From Symbolic to Geometric: Enabling Spatial Reasoning in Large Language Models

DGX agent

arXiv:2606.04381v1 Announce Type: cross Abstract: Recent large language models (LLMs) often appear to exhibit spatial reasoning ability; however, this capability is largely symbolic, arising from patt

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

From Untrusted Input to Trusted Memory: A Systematic Study of Memory Poisoning Attacks in LLM Agents

DGX agent

arXiv:2606.04329v1 Announce Type: cross Abstract: Memory is a core component of AI agents, enabling them to accumulate knowledge across interactions and improve performance. However, persistent memory

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

GeM-NR: Geometry-Aware Multi-View Editing for Nonrigid Scene Changes

DGX agent

arXiv:2606.05142v1 Announce Type: cross Abstract: Recent developments in multi-view image editing with generative models have brought us a step closer toward general 3D content generation and customiz

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

GENEB: Why Genomic Models Are Hard to Compare

DGX agent

arXiv:2606.04525v1 Announce Type: new Abstract: Progress in genomic foundation models is difficult to assess due to fragmented benchmarks, incompatible evaluation protocols, and task-specific reportin

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

Geometry-Aware Hallucination Detection in Large Language Models

DGX agent

arXiv:2601.06196v3 Announce Type: replace-cross Abstract: Large language models (LLMs) frequently generate factually incorrect or unsupported content, commonly referred to as hallucinations. Prior wor

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Geometry Gaussians: Decoupling Appearance and Geometry in Gaussian Splatting

DGX agent

arXiv:2606.05124v1 Announce Type: cross Abstract: After the success of 3D Gaussian Splatting (3DGS) for novel view synthesis, many works have explored how to also use it for geometric surface represen

model-releasesarxiv-cs-cv
4 Jun 2026
← Previous
1…214215216217218…475
Next →