AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
All
85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,778 results
Model Releases

Robust and Efficient Guardrails with Latent Reasoning

DGX agent

arXiv:2605.29068v1 Announce Type: new Abstract: Maintaining the safety of large language models (LLMs) is crucial as they are increasingly deployed in real-world applications. Existing safety guardrai

model-releasesarxiv-cs-ai
29 May 2026
Model Releases
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

S-MARC: Causal Streaming Reasoning for Full-Duplex Conversational Behavior Modeling

DGX agent

arXiv:2602.11065v2 Announce Type: replace-cross Abstract: Human conversation is organized by an implicit chain of thought and manifests as temporally structured conversational behaviors. Capturing thi

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

SAAS: Self-Aware Reinforcement Learning for Over-Search Mitigation in Agentic Search

DGX agent

arXiv:2605.29796v1 Announce Type: new Abstract: Agentic search enables LLMs to solve complex multi-hop questions through iterative reasoning and external search. Despite the effectiveness, these syste

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

SafeSearch: Automated Red-Teaming of LLM-Based Search Agents

DGX agent

arXiv:2509.23694v5 Announce Type: replace Abstract: Search agents connect LLMs to the Internet, enabling them to access broader and more up-to-date information. However, this also introduces a new thr

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

SAGE: Segment-Aware Gloss-Free Encoding for Token-Efficient Sign Language Translation

DGX agent

arXiv:2507.09266v2 Announce Type: replace Abstract: Gloss-free Sign Language Translation (SLT) has advanced rapidly, achieving strong performances without relying on gloss annotations. However, these

model-releasesarxiv-cs-cv
29 May 2026
Model Releases

Salesforce published a detailed writeup on going agentic with Claude Code. A couple things jumped out. A migration they'd scoped at 231 days…

DGX agent

Salesforce published a detailed writeup on going agentic with Claude Code. A couple things jumped out. A migration they'd scoped at 231 days shipped in 13. One PR delivered 21 endpoints at 100% test c

model-releasesboris-cherny--x
29 May 2026
Model Releases

Same Question, Different Source, Different Answer: Auditing Source-Dependence in Medical Multi-Source RAG

DGX agent

arXiv:2605.29084v1 Announce Type: cross Abstract: A retrieval-augmented generation (RAG) system deployed over a multi-author institutional corpus can give a different answer to the same question depen

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Sample-Efficient Diffusion-based Reinforcement Learning with Critic Guidance

DGX agent

arXiv:2605.30056v1 Announce Type: cross Abstract: Recent advances in reinforcement learning (RL) have achieved great successes by leveraging the multimodality and exploration capability of diffusion p

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

Scaling Laws for Agent Harnesses via Effective Feedback Compute

DGX agent

arXiv:2605.29682v1 Announce Type: new Abstract: Agent harnesses increasingly determine the performance of language-model systems by deciding how models call tools, receive feedback, verify intermediat

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

Scaling Monosemanticity: Extracting Interpretable Features from Claude 3 Sonnet

DGX agent

arXiv:2605.29358v1 Announce Type: new Abstract: We demonstrate that sparse autoencoders can extract interpretable features from Claude 3 Sonnet, a production-scale language model, addressing the open

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

SCDBench: A Benchmark for LLM-Based Smart Contract Decompilers

DGX agent

arXiv:2605.29059v1 Announce Type: cross Abstract: Smart contract decompilation aims to recover high-level source code from bytecode, but evaluating decompilers remains difficult because existing studi

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

SciIntBench: Measuring LLM Compliance with Research Integrity Norms Under Adversarial Framing

DGX agent

arXiv:2605.29468v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to support scientific work, but it is unclear whether they uphold responsible conduct of research (

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

SCOPE: Prompt Evolution for Enhancing Agent Effectiveness

DGX agent

arXiv:2512.15374v2 Announce Type: replace Abstract: Large Language Model (LLM) agents are increasingly deployed in environments that generate massive, dynamic contexts. However, a critical bottleneck

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

SDF-Net: Structure-Aware Disentangled Feature Learning for Opticall-SAR Ship Re-identification

DGX agent

arXiv:2603.12588v2 Announce Type: replace Abstract: Cross-modal ship re-identification (ReID) between optical and synthetic aperture radar (SAR) imagery is fundamentally challenged by the severe radio

model-releasesarxiv-cs-cv
29 May 2026
Model Releases

Selection Hyper-heuristics Can Automatically Adjust the Learning Period to Optimally Solve Pseudo-Boolean Problems

DGX agent

arXiv:2605.29916v1 Announce Type: cross Abstract: The Random Gradient hyper-heuristic was recently shown to be able to learn the optimal neighbourhood size when optimizing the LeadingOnes benchmark vi

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Selective QA over Conflicting Multi-Source Personal Memory: A Diagnostic Testbed and Method Comparison

DGX agent

arXiv:2605.30087v1 Announce Type: new Abstract: Emerging personal AI agents are moving toward persistent, multi-source memory. This creates an evaluation problem: systems must decide how to use confli

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Semantic and Visual Evidence for Efficient Long-Video Reasoning: A Solution for the HD-EPIC VQA Challenge

DGX agent

arXiv:2605.29402v1 Announce Type: cross Abstract: Understanding long-form egocentric videos remains challenging for multimodal large language models (MLLMs) due to limited context length and insuffici

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Sequential Physics-Constrained Neural Operator Forward Modeling for the extit{Norne} Reservoir System

DGX agent

arXiv:2605.28909v1 Announce Type: new Abstract: We develop a comprehensive mathematical and computational framework for sequential surrogate modeling of three-phase black-oil reservoir dynamics using

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

SERC: LDPC-Inspired Semantic Error Correction for Retrieval-Augmented Generation

DGX agent

arXiv:2605.28837v1 Announce Type: cross Abstract: While Large Language Models (LLMs) have demonstrated remarkable capabilities, their reliability is significantly compromised by hallucinations. Existi

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Skill-Pro: Learning Reusable Skills from Experience via Non-Parametric PPO for LLM Agents

DGX agent

arXiv:2602.01869v3 Announce Type: replace Abstract: LLM-driven agents excel at sequential decision-making but often rely on on-the-fly reasoning, re-deriving solutions even in recurring scenarios. Thi

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

SkillsInjector: Dynamic Skill Context Construction for LLM Agents

DGX agent

arXiv:2605.29794v1 Announce Type: new Abstract: LLM agents now draw on growing skill libraries to handle complex tasks. However, injecting more skills does not always improve task completion and can e

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

SLAD : Shared LoRA Adapters for Task Specific Distillation

DGX agent

arXiv:2605.29726v1 Announce Type: new Abstract: In the context of resource-constrained environments such as embedded systems, adapting reduced-size foundation models to downstream tasks has become inc

model-releasesarxiv-cs-cv
29 May 2026
Model Releases

Small Agent Group is the Future of Digital Health

DGX agent

arXiv:2602.08013v2 Announce Type: replace Abstract: The rapid adoption of large language models (LLMs) in digital health has been driven by a 'scaling-first' philosophy, i.e., the assumption that clin

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

SMolLM: Small Language Models Learn Small Molecular Grammar

DGX agent

arXiv:2605.06322v2 Announce Type: replace Abstract: Language models for molecular design have scaled to hundreds of millions of parameters, yet how they learn chemical grammar is poorly understood. We

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

Some fun Gemini Omni use cases from the community 🧵👇

DGX agent

This X thread from Google AI showcases community-created use cases and applications of Gemini Omni, Google's multimodal AI model. The post likely highlights practical and creative examples of how user

model-releasesgoogle-ai--x
29 May 2026
Model Releases

SoundnessBench: Can Your AI Scientist Really Tell Good Research Ideas from Bad Ones?

DGX agent

arXiv:2605.30329v1 Announce Type: new Abstract: Autonomous AI research agents aim to accelerate scientific discovery by automating the research pipeline, from hypothesis generation to peer review. How

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

Stable-Layers: Fine-Tuning Image Layer Decomposition Models with VLM-Scored Reinforcement Learning

DGX agent

arXiv:2605.30257v1 Announce Type: new Abstract: We present Stable-Layers, a reinforcement learning framework that eliminates the need for paired supervision by fine-tuning a pretrained layer decomposi

model-releasesarxiv-cs-cv
29 May 2026
Model Releases

STAMP: Training Explicit Memory for Mobile GUI Agents in Controllable and Scalable Virtual Environments

DGX agent

arXiv:2605.29324v1 Announce Type: new Abstract: Mobile GUI agents excel at immediate reactive control but frequently fail in realistic, long-horizon tasks that require memory. This failure stems from

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

Strengthening societal resilience with Rosalind Biodefense

DGX agent

OpenAI launches Rosalind Biodefense, expanding trusted access to GPT-Rosalind for vetted developers and U.S. government partners advancing biodefense, public health, and pandemic preparedness through

model-releasesopenai
29 May 2026
Model Releases

Striding Across Reynolds Numbers: Representation Geometry in Neural PDE Generalisation

DGX agent

arXiv:2605.30112v1 Announce Type: new Abstract: Cross-Reynolds generalisation in neural PDE solvers remains poorly characterised. On the canonical forced 2D Navier-Stokes benchmark, a trained Fourier

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

Structure-Aware Text Recognition for Ancient Greek Critical Editions

DGX agent

arXiv:2603.02803v2 Announce Type: replace Abstract: Recent advances in visual language models (VLMs) have transformed end-to-end document understanding. However, their ability to interpret the complex

model-releasesarxiv-cs-cv
29 May 2026
Model Releases

SURGENT: A Surgical Multi-Agent Assistance System Across the Perioperative Workflow

DGX agent

arXiv:2605.29368v1 Announce Type: cross Abstract: The intricate nature of modern surgical care necessitates intelligent systems that can synthesize extensive patient records, support collaborative dec

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

SwInception -- Local Attention Meets Convolutions

DGX agent

arXiv:2605.29954v1 Announce Type: new Abstract: Sparse vision transformers have gained popularity as efficient encoders for medical volumetric segmentation, with Swin emerging as a prominent choice. S

model-releasesarxiv-cs-cv
29 May 2026
Model Releases

TAE: Target-aware enhancer for nighttime UAV tracking

DGX agent

arXiv:2605.29558v1 Announce Type: new Abstract: Severe image degradation under low-light nighttime conditions constitutes a core bottleneck preventing all-day applications for UAV-based single object

model-releasesarxiv-cs-cv
29 May 2026
Model Releases

TANDEM: Temporal-Aware Neural Detection for Multimodal Hate Speech

DGX agent

arXiv:2601.11178v2 Announce Type: replace Abstract: Social media platforms are increasingly dominated by long-form multimodal content, where harmful narratives are constructed through a complex interp

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

TaxDistill: Improving Metagenomic Taxonomic Annotation via Distilled Genomic Foundation Models

DGX agent

arXiv:2605.28868v1 Announce Type: cross Abstract: Metagenomic taxonomic annotation aims to identify the microbial origins of DNA fragments in environmental samples. Traditional methods that rely on se

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Teaching Language Models to Check Grounded Claim Factuality with Human Test-Taking Strategies

DGX agent

arXiv:2605.29712v1 Announce Type: cross Abstract: Grounded claim factuality checking is important for large language model (LLM) applications such as retrieval-augmented generation, as it helps users

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Temporal Stability and Few-Shot Prompting in Math Task Assessment

DGX agent

arXiv:2605.30151v1 Announce Type: new Abstract: As AI tools become increasingly integrated into educational contexts, questions arise about both their stability over time and their responsiveness to p

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Text-Preserving Lossy Text Compression: A Study of Strategic Deletion and LLM Reconstruction

DGX agent

arXiv:2605.29000v1 Announce Type: new Abstract: Traditional lossless text compression preserves every byte, but its gains on natural language are often modest in realistic operating regimes. We study

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

The blast bent those steel beams on the tower inwards:

DGX agent

The blast bent those steel beams on the tower inwards: First look at LC-36 from the air this morning after the explosion of New Glenn last night during a failed hotfire test. Visible is the wreckage f

model-releasesanthropic--x
29 May 2026
Model Releases

The Chain Holds, the Answer Folds: Trace-Answer Dissociation in Reasoning Models Under Adversarial Pressure

DGX agent

arXiv:2605.29087v1 Announce Type: new Abstract: Reasoning models are evaluated on single-turn benchmarks but deployed in multi-turn dialogue, where users push back on correct answers. Under sustained

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

The Cognitive Categorical Transformer: Category-Theoretic Inductive Biases for Language Modeling

DGX agent

arXiv:2605.28864v1 Announce Type: new Abstract: The Cognitive Categorical Transformer (CCT) is a 306M-parameter architecture that augments a pretrained GPT-2 Small backbone with cognitively grounded c

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

The Curse of Helpfulness: Inverse Scaling Law in Robustness to Distractor Instructions via DistractionIF

DGX agent

arXiv:2605.29491v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed in agentic and retrieval-augmented generation (RAG) systems, where they must execute user-specifi

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

The Good, the Bad, and the Ugly of Markov Boundary for Tabular Prediction

DGX agent

arXiv:2605.29411v1 Announce Type: cross Abstract: Under standard graphical assumptions, the Markov boundary of a target variable is the smallest set of features that renders every other feature redund

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

The Hamilton-Jacobi Theory of Deep Learning

DGX agent

arXiv:2605.28983v1 Announce Type: cross Abstract: In this paper, training a neural network is identified, exactly, as a search through Hamilton--Jacobi initial-value problems: each gradient step selec

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

The improvements run wide. Across all major European languages, Command A+ consistently pulls ahead of competitors on WMT24++ (xCOMET-XL): …

DGX agent

The improvements run wide. Across all major European languages, Command A+ consistently pulls ahead of competitors on WMT24++ (xCOMET-XL): 🇫🇷 +2.4 pts in French 🇪🇸 +1.9 pts in Spanish 🇩🇪 +0.9 pts in G

model-releasescohere--x
29 May 2026
Model Releases

The Open Motion Planning Library 2.0

DGX agent

arXiv:2605.29301v1 Announce Type: new Abstract: The Open Motion Planning Library (OMPL), first released in 2008, has become a cornerstone of the motion planning community, providing implementations of

model-releasesarxiv-cs-ro
29 May 2026
Model Releases

The Price Reversal Phenomenon: When Cheaper Reasoning Models Cost More

DGX agent

arXiv:2603.23971v2 Announce Type: replace-cross Abstract: Developers and consumers increasingly choose reasoning models (RMs) based on their listed API prices. However, how accurately do these prices

model-releasesarxiv-cs-ai
29 May 2026
← Previous
1…249250251252253…475
Next →