AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
All
85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,770 results
Model Releases

Scaling Laws for Agent Harnesses via Effective Feedback Compute

DGX agent

arXiv:2605.29682v1 Announce Type: new Abstract: Agent harnesses increasingly determine the performance of language-model systems by deciding how models call tools, receive feedback, verify intermediat

model-releasesarxiv-cs-cl
29 May 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Scaling Monosemanticity: Extracting Interpretable Features from Claude 3 Sonnet

DGX agent

arXiv:2605.29358v1 Announce Type: new Abstract: We demonstrate that sparse autoencoders can extract interpretable features from Claude 3 Sonnet, a production-scale language model, addressing the open

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

SCDBench: A Benchmark for LLM-Based Smart Contract Decompilers

DGX agent

arXiv:2605.29059v1 Announce Type: cross Abstract: Smart contract decompilation aims to recover high-level source code from bytecode, but evaluating decompilers remains difficult because existing studi

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

SciIntBench: Measuring LLM Compliance with Research Integrity Norms Under Adversarial Framing

DGX agent

arXiv:2605.29468v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to support scientific work, but it is unclear whether they uphold responsible conduct of research (

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

SCOPE: Prompt Evolution for Enhancing Agent Effectiveness

DGX agent

arXiv:2512.15374v2 Announce Type: replace Abstract: Large Language Model (LLM) agents are increasingly deployed in environments that generate massive, dynamic contexts. However, a critical bottleneck

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

SDF-Net: Structure-Aware Disentangled Feature Learning for Opticall-SAR Ship Re-identification

DGX agent

arXiv:2603.12588v2 Announce Type: replace Abstract: Cross-modal ship re-identification (ReID) between optical and synthetic aperture radar (SAR) imagery is fundamentally challenged by the severe radio

model-releasesarxiv-cs-cv
29 May 2026
Model Releases

Selection Hyper-heuristics Can Automatically Adjust the Learning Period to Optimally Solve Pseudo-Boolean Problems

DGX agent

arXiv:2605.29916v1 Announce Type: cross Abstract: The Random Gradient hyper-heuristic was recently shown to be able to learn the optimal neighbourhood size when optimizing the LeadingOnes benchmark vi

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Selective QA over Conflicting Multi-Source Personal Memory: A Diagnostic Testbed and Method Comparison

DGX agent

arXiv:2605.30087v1 Announce Type: new Abstract: Emerging personal AI agents are moving toward persistent, multi-source memory. This creates an evaluation problem: systems must decide how to use confli

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Semantic and Visual Evidence for Efficient Long-Video Reasoning: A Solution for the HD-EPIC VQA Challenge

DGX agent

arXiv:2605.29402v1 Announce Type: cross Abstract: Understanding long-form egocentric videos remains challenging for multimodal large language models (MLLMs) due to limited context length and insuffici

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Sequential Physics-Constrained Neural Operator Forward Modeling for the extit{Norne} Reservoir System

DGX agent

arXiv:2605.28909v1 Announce Type: new Abstract: We develop a comprehensive mathematical and computational framework for sequential surrogate modeling of three-phase black-oil reservoir dynamics using

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

SERC: LDPC-Inspired Semantic Error Correction for Retrieval-Augmented Generation

DGX agent

arXiv:2605.28837v1 Announce Type: cross Abstract: While Large Language Models (LLMs) have demonstrated remarkable capabilities, their reliability is significantly compromised by hallucinations. Existi

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

DGX agent

arXiv:2602.01869v3 Announce Type: replace Abstract: LLM-driven agents excel at sequential decision-making but often rely on on-the-fly reasoning, re-deriving solutions even in recurring scenarios. Thi

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

SkillsInjector: Dynamic Skill Context Construction for LLM Agents

DGX agent

arXiv:2605.29794v1 Announce Type: new Abstract: LLM agents now draw on growing skill libraries to handle complex tasks. However, injecting more skills does not always improve task completion and can e

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

SLAD : Shared LoRA Adapters for Task Specific Distillation

DGX agent

arXiv:2605.29726v1 Announce Type: new Abstract: In the context of resource-constrained environments such as embedded systems, adapting reduced-size foundation models to downstream tasks has become inc

model-releasesarxiv-cs-cv
29 May 2026
Model Releases

Small Agent Group is the Future of Digital Health

DGX agent

arXiv:2602.08013v2 Announce Type: replace Abstract: The rapid adoption of large language models (LLMs) in digital health has been driven by a 'scaling-first' philosophy, i.e., the assumption that clin

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

SMolLM: Small Language Models Learn Small Molecular Grammar

DGX agent

arXiv:2605.06322v2 Announce Type: replace Abstract: Language models for molecular design have scaled to hundreds of millions of parameters, yet how they learn chemical grammar is poorly understood. We

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

Some fun Gemini Omni use cases from the community 🧵👇

DGX agent

This X thread from Google AI showcases community-created use cases and applications of Gemini Omni, Google's multimodal AI model. The post likely highlights practical and creative examples of how user

model-releasesgoogle-ai--x
29 May 2026
Model Releases

SoundnessBench: Can Your AI Scientist Really Tell Good Research Ideas from Bad Ones?

DGX agent

arXiv:2605.30329v1 Announce Type: new Abstract: Autonomous AI research agents aim to accelerate scientific discovery by automating the research pipeline, from hypothesis generation to peer review. How

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

Stable-Layers: Fine-Tuning Image Layer Decomposition Models with VLM-Scored Reinforcement Learning

DGX agent

arXiv:2605.30257v1 Announce Type: new Abstract: We present Stable-Layers, a reinforcement learning framework that eliminates the need for paired supervision by fine-tuning a pretrained layer decomposi

model-releasesarxiv-cs-cv
29 May 2026
Model Releases

STAMP: Training Explicit Memory for Mobile GUI Agents in Controllable and Scalable Virtual Environments

DGX agent

arXiv:2605.29324v1 Announce Type: new Abstract: Mobile GUI agents excel at immediate reactive control but frequently fail in realistic, long-horizon tasks that require memory. This failure stems from

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

Strengthening societal resilience with Rosalind Biodefense

DGX agent

OpenAI launches Rosalind Biodefense, expanding trusted access to GPT-Rosalind for vetted developers and U.S. government partners advancing biodefense, public health, and pandemic preparedness through

model-releasesopenai
29 May 2026
Model Releases

Striding Across Reynolds Numbers: Representation Geometry in Neural PDE Generalisation

DGX agent

arXiv:2605.30112v1 Announce Type: new Abstract: Cross-Reynolds generalisation in neural PDE solvers remains poorly characterised. On the canonical forced 2D Navier-Stokes benchmark, a trained Fourier

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

Structure-Aware Text Recognition for Ancient Greek Critical Editions

DGX agent

arXiv:2603.02803v2 Announce Type: replace Abstract: Recent advances in visual language models (VLMs) have transformed end-to-end document understanding. However, their ability to interpret the complex

model-releasesarxiv-cs-cv
29 May 2026
Model Releases

SURGENT: A Surgical Multi-Agent Assistance System Across the Perioperative Workflow

DGX agent

arXiv:2605.29368v1 Announce Type: cross Abstract: The intricate nature of modern surgical care necessitates intelligent systems that can synthesize extensive patient records, support collaborative dec

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

SwInception -- Local Attention Meets Convolutions

DGX agent

arXiv:2605.29954v1 Announce Type: new Abstract: Sparse vision transformers have gained popularity as efficient encoders for medical volumetric segmentation, with Swin emerging as a prominent choice. S

model-releasesarxiv-cs-cv
29 May 2026
Model Releases

TAE: Target-aware enhancer for nighttime UAV tracking

DGX agent

arXiv:2605.29558v1 Announce Type: new Abstract: Severe image degradation under low-light nighttime conditions constitutes a core bottleneck preventing all-day applications for UAV-based single object

model-releasesarxiv-cs-cv
29 May 2026
Model Releases

TANDEM: Temporal-Aware Neural Detection for Multimodal Hate Speech

DGX agent

arXiv:2601.11178v2 Announce Type: replace Abstract: Social media platforms are increasingly dominated by long-form multimodal content, where harmful narratives are constructed through a complex interp

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

TaxDistill: Improving Metagenomic Taxonomic Annotation via Distilled Genomic Foundation Models

DGX agent

arXiv:2605.28868v1 Announce Type: cross Abstract: Metagenomic taxonomic annotation aims to identify the microbial origins of DNA fragments in environmental samples. Traditional methods that rely on se

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Teaching Language Models to Check Grounded Claim Factuality with Human Test-Taking Strategies

DGX agent

arXiv:2605.29712v1 Announce Type: cross Abstract: Grounded claim factuality checking is important for large language model (LLM) applications such as retrieval-augmented generation, as it helps users

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Temporal Stability and Few-Shot Prompting in Math Task Assessment

DGX agent

arXiv:2605.30151v1 Announce Type: new Abstract: As AI tools become increasingly integrated into educational contexts, questions arise about both their stability over time and their responsiveness to p

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Text-Preserving Lossy Text Compression: A Study of Strategic Deletion and LLM Reconstruction

DGX agent

arXiv:2605.29000v1 Announce Type: new Abstract: Traditional lossless text compression preserves every byte, but its gains on natural language are often modest in realistic operating regimes. We study

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

The blast bent those steel beams on the tower inwards:

DGX agent

The blast bent those steel beams on the tower inwards: First look at LC-36 from the air this morning after the explosion of New Glenn last night during a failed hotfire test. Visible is the wreckage f

model-releasesanthropic--x
29 May 2026
Model Releases

The Chain Holds, the Answer Folds: Trace-Answer Dissociation in Reasoning Models Under Adversarial Pressure

DGX agent

arXiv:2605.29087v1 Announce Type: new Abstract: Reasoning models are evaluated on single-turn benchmarks but deployed in multi-turn dialogue, where users push back on correct answers. Under sustained

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

The Cognitive Categorical Transformer: Category-Theoretic Inductive Biases for Language Modeling

DGX agent

arXiv:2605.28864v1 Announce Type: new Abstract: The Cognitive Categorical Transformer (CCT) is a 306M-parameter architecture that augments a pretrained GPT-2 Small backbone with cognitively grounded c

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

The Curse of Helpfulness: Inverse Scaling Law in Robustness to Distractor Instructions via DistractionIF

DGX agent

arXiv:2605.29491v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed in agentic and retrieval-augmented generation (RAG) systems, where they must execute user-specifi

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

The Good, the Bad, and the Ugly of Markov Boundary for Tabular Prediction

DGX agent

arXiv:2605.29411v1 Announce Type: cross Abstract: Under standard graphical assumptions, the Markov boundary of a target variable is the smallest set of features that renders every other feature redund

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

The Hamilton-Jacobi Theory of Deep Learning

DGX agent

arXiv:2605.28983v1 Announce Type: cross Abstract: In this paper, training a neural network is identified, exactly, as a search through Hamilton--Jacobi initial-value problems: each gradient step selec

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

The improvements run wide. Across all major European languages, Command A+ consistently pulls ahead of competitors on WMT24++ (xCOMET-XL): …

DGX agent

The improvements run wide. Across all major European languages, Command A+ consistently pulls ahead of competitors on WMT24++ (xCOMET-XL): 🇫🇷 +2.4 pts in French 🇪🇸 +1.9 pts in Spanish 🇩🇪 +0.9 pts in G

model-releasescohere--x
29 May 2026
Model Releases

The Open Motion Planning Library 2.0

DGX agent

arXiv:2605.29301v1 Announce Type: new Abstract: The Open Motion Planning Library (OMPL), first released in 2008, has become a cornerstone of the motion planning community, providing implementations of

model-releasesarxiv-cs-ro
29 May 2026
Model Releases

The Price Reversal Phenomenon: When Cheaper Reasoning Models Cost More

DGX agent

arXiv:2603.23971v2 Announce Type: replace-cross Abstract: Developers and consumers increasingly choose reasoning models (RMs) based on their listed API prices. However, how accurately do these prices

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

The story gets bigger beyond Europe. Command A+ makes major gains in high-impact non-Latin languages; outperforming Mistral Medium 3.5 in Ko…

DGX agent

The story gets bigger beyond Europe. Command A+ makes major gains in high-impact non-Latin languages; outperforming Mistral Medium 3.5 in Korean, Japanese, Hebrew, Chinese, and Arabic. For Arabic, tha

model-releasescohere--x
29 May 2026
Model Releases

The team at @llama_index built an awesome template using LlamaParse and the new Managed Agents in the Gemini API. See how they built an agen…

DGX agent

The team at @llama_index built an awesome template using LlamaParse and the new Managed Agents in the Gemini API. See how they built an agent that can tackle unstructured documents. 📄↓ 🚀 The team at @

model-releasesjerry-liu--x
29 May 2026
Model Releases

The Trust Paradox: How CS Researchers Engage LLM Leaderboards

DGX agent

arXiv:2605.28966v1 Announce Type: new Abstract: Large language model (LLM) leaderboards rank AI models using standardized benchmarks and have become highly visible across computer science, despite kno

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

The Vision Wormhole: Latent-Space Communication in Heterogeneous Multi-Agent Systems

DGX agent

arXiv:2602.15382v2 Announce Type: replace Abstract: Multi-Agent Systems (MAS) powered by Large Language Models have unlocked advanced collaborative reasoning, yet they remain bottlenecked by discrete

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

This is a diary entry to myself, so I remember what AI was like today. It's just going to be a bullet-list stream of consciousness. - There …

DGX agent

This is a diary entry to myself, so I remember what AI was like today. It's just going to be a bullet-list stream of consciousness. - There are still so many leaders that have never seen an agent run

model-releasesallie-k--miller--x
29 May 2026
Model Releases

this was my pypi hardening strategy for @activegraphai: - spin up a new @replit - point at docs page, ask it to build something - ask it to …

DGX agent

this was my pypi hardening strategy for @activegraphai: - spin up a new @replit - point at docs page, ask it to build something - ask it to write a feedback report to package builder - feed that feedb

model-releasesyohei-nakajima--x
29 May 2026
Model Releases

Three-dimensional Conditional Diffusion Models for Cosmological 21 cm Lightcone Emulation

DGX agent

arXiv:2605.29016v1 Announce Type: cross Abstract: We investigate conditional diffusion modeling for three-dimensional 21 cm lightcone emulation, focusing on cubes with a sky-plane size of 64imes64 and

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

TIMEGATE: Sustainable Time-Boxed Promotion Gates for Continual ML Adaptation Under Resource Constraints

DGX agent

arXiv:2605.29183v1 Announce Type: cross Abstract: As machine learning(ML) systems evolve to continual adaptation, each re-training cycle uses compute, annotation, and energy. We introduce TIMEGATE, a

model-releasesarxiv-cs-ai
29 May 2026
← Previous
1…248249250251252…475
Next →