AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

SciResearcher: Scaling Deep Research Agents for Frontier Scientific Reasoning

DGX agent

arXiv:2605.01489v1 Announce Type: cross Abstract: Frontier scientific reasoning is rapidly emerging as a key foundation for advancing AI agents in automated scientific discovery. Deep research agents

model-releasesarxiv-cs-cl
5 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Seeing Realism from Simulation: Efficient Video Transfer for Vision-Language-Action Data Augmentation

DGX agent

arXiv:2605.02757v1 Announce Type: new Abstract: Vision-language-action (VLA) models typically rely on large-scale real-world videos, whereas simulated data, despite being inexpensive and highly parall

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Seeing the Scene Matters: Revealing Forgetting in Video Understanding Models with a Scene-Aware Long-Video Benchmark

DGX agent

arXiv:2603.27259v2 Announce Type: replace Abstract: Long video understanding (LVU) remains a core challenge in multimodal learning. Although recent vision-language models (VLMs) have made notable prog

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Segment-Aligned Policy Optimization for Multi-Modal Reasoning

DGX agent

arXiv:2605.01327v1 Announce Type: cross Abstract: Existing reinforcement learning approaches for Large Language Models typically perform policy optimization at the granularity of individual tokens or

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Selector-Guided Autonomous Curriculum for One-Shot Reinforcement Learning from Verifiable Rewards

DGX agent

arXiv:2605.01823v1 Announce Type: new Abstract: Recently, Reinforcement Learning from Verifiable Rewards (RLVR) has been established as a highly effective technique for augmenting the math reasoning s

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Self-Supervised Learning for Multimodal Non-Rigid 3D Shape Matching

DGX agent

arXiv:2303.10971v2 Announce Type: replace Abstract: The matching of 3D shapes has been extensively studied for shapes represented as surface meshes, as well as for shapes represented as point clouds.

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

SemEval-2026 Task 7: Everyday Knowledge Across Diverse Languages and Cultures

DGX agent

arXiv:2605.02601v1 Announce Type: new Abstract: We present our shared task on evaluating the adaptability of LLMs and NLP systems across multiple languages and cultures. The task data consist of an ex

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Sentinel-VLA: A Metacognitive VLA Model with Active Status Monitoring for Dynamic Reasoning and Error Recovery

DGX agent

arXiv:2605.01191v1 Announce Type: new Abstract: Vision-language-action (VLA) models have advanced the field of embodied manipulation by harnessing broad world knowledge and strong generalization. Howe

model-releasesarxiv-cs-ro
5 May 2026
Model Releases

SF20K Competition 2025: Summary and findings

DGX agent

arXiv:2605.01496v1 Announce Type: new Abstract: This report presents the results and findings of the first edition of the Short-Films 20K (SF20K) Competition, held in conjunction with the SLoMO Worksh

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Shadow-Loom: Causal Reasoning over Graphical World Model of Narratives

DGX agent

arXiv:2605.02475v1 Announce Type: cross Abstract: Stories hold a reader's attention because they have causes, secrets, and consequences. Shadow-Loom is an experimental open-source framework that turns

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Sharpness-Aware Pretraining Mitigates Catastrophic Forgetting

DGX agent

arXiv:2605.02105v1 Announce Type: cross Abstract: Pretraining optimizers are tuned to produce the strongest possible base model, on the assumption that a stronger starting point yields a stronger mode

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Silicon Showdown: Performance, Efficiency, and Ecosystem Barriers in Consumer-Grade LLM Inference

DGX agent

arXiv:2605.00519v2 Announce Type: cross Abstract: The operational landscape of local Large Language Model (LLM) inference has shifted from lightweight models to datacenter-class weights exceeding 70B

model-releasesarxiv-cs-ai
5 May 2026
Model Releases

Social Bias in LLM-Generated Code: Benchmark and Mitigation

DGX agent

arXiv:2605.00382v2 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly deployed to generate code for human-centered applications where demographic fairness is critical. Howeve

model-releasesarxiv-cs-ai
5 May 2026
Model Releases

Sparse Regression under Correlation and Weak Signals: A Reproducible Benchmark of Classical and Bayesian Methods

DGX agent

arXiv:2605.00835v1 Announce Type: new Abstract: Choosing between classical and Bayesian sparse regression methods involves a real trade-off: penalized estimators like Lasso run in milliseconds but giv

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

SpecEdit: Training-Free Acceleration for Diffusion based Image Editing via Semantic Locking

DGX agent

arXiv:2605.02152v1 Announce Type: new Abstract: Diffusion-based image editing offers strong semantic controllability, but remains computationally expensive due to iterative high-resolution denoising o

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Spectral Model eXplainer: a chemically-grounded explainability framework for spectral-based machine learning models

DGX agent

arXiv:2605.02684v1 Announce Type: new Abstract: Spectral-based machine learning models have been increasingly deployed in chemometrics and spectroscopy, where predictive accuracy is as important as ex

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

SplAttN: Bridging 2D and 3D with Gaussian Soft Splatting and Attention for Point Cloud Completion

DGX agent

arXiv:2605.01466v1 Announce Type: new Abstract: Although multi-modal learning has advanced point cloud completion, the theoretical mechanisms remain unclear. Recent works attribute success to the conn

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Split-on-Share: Mixture of Sparse Experts for Task-Agnostic Continual Learning

DGX agent

arXiv:2601.17616v2 Announce Type: replace Abstract: Continual learning in Large Language Models (LLMs) is hindered by the plasticity-stability dilemma, where acquiring new capabilities often leads to

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Spoken Language Identification with Pre-trained Models and Margin Loss

DGX agent

arXiv:2605.01905v1 Announce Type: cross Abstract: For the speaker-controlled spoken language identification task proposed in the TidyLang Challenge 2026, this paper proposes a language identification

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

SRTJ: Self-Evolving Rule-Driven Training-Free LLM Jailbreaking

DGX agent

arXiv:2605.00974v1 Announce Type: cross Abstract: LLMs are increasingly equipped with safety alignment mechanisms, yet recent studies demonstrate that they remain vulnerable to jailbreaking attacks th

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

STAGE: A Full-Screenplay Benchmark for Reasoning over Evolving Storie

DGX agent

arXiv:2601.08510v3 Announce Type: replace Abstract: Movie screenplays are rich long-form narratives that interleave complex character relationships, temporally ordered events, and dialogue-driven inte

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Standing on the Shoulders of Giants: Stabilized Knowledge Distillation for Cross--Language Code Clone Detection

DGX agent

arXiv:2605.02860v1 Announce Type: cross Abstract: Cross-language code clone detection (X-CCD) is challenging because semantically equivalent programs written in different languages often share little

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Statistically-Lossless Quantization of Large Language Models

DGX agent

arXiv:2605.02404v1 Announce Type: new Abstract: Model quantization has become essential for efficient large language model deployment, yet existing approaches involve clear trade-offs: methods such as

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

SteeringDiffusion: A Bottlenecked Activation Control Interface for Diffusion Models

DGX agent

arXiv:2605.01653v1 Announce Type: new Abstract: We introduce SteeringDiffusion, a bottlenecked activation-level control interface for diffusion models that exposes a smooth, monotonic, and runtime-adj

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

STEP: Warm-Started Visuomotor Policies with Spatiotemporal Consistency Prediction

DGX agent

arXiv:2602.08245v2 Announce Type: replace Abstract: Diffusion policies have recently emerged as a powerful paradigm for visuomotor control in robotic manipulation due to their ability to model the dis

model-releasesarxiv-cs-ro
5 May 2026
Model Releases

StereoMamba: Real-time and Robust Intraoperative Stereo Disparity Estimation via Long-range Spatial Dependencies

DGX agent

arXiv:2504.17401v2 Announce Type: replace Abstract: Stereo disparity estimation is crucial for obtaining depth information in robot-assisted minimally invasive surgery (RAMIS). While current deep lear

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

StreamIndex: Memory-Bounded Compressed Sparse Attention via Streaming Top-k

DGX agent

arXiv:2605.02568v1 Announce Type: new Abstract: DeepSeek-V3.2 and V4 introduce Compressed Sparse Attention (CSA): a lightning indexer (a learned scoring projection over compressed keys) scores them, t

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

StressEval: Failure-Driven Dynamic Benchmarking for Knowledge-Intensive Reasoning in Large Language Models

DGX agent

arXiv:2605.01939v1 Announce Type: new Abstract: Static benchmarks for LLMs are increasingly compromised by contamination and overfitting especially on knowledge intensive reasoning tasks While recent

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

StyleShield: Exposing the Fragility of AIGC Detectors through Continuous Controllable Style Transfer

DGX agent

arXiv:2605.00924v1 Announce Type: new Abstract: AI-generated content (AIGC) detectors are increasingly deployed in high-stakes settings such as academic integrity screening, yet their reliability rest

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Submodular Benchmark Selection

DGX agent

arXiv:2605.02209v1 Announce Type: cross Abstract: Evaluating large language models across many benchmarks is expensive, yet many benchmarks are highly correlated. We formalize the selection of a small

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

SurgCheck: Do Vision-Language Models Really Look at Images in Surgical VQA?

DGX agent

arXiv:2605.01911v1 Announce Type: new Abstract: Purpose: Vision-language models (VLMs) have shown promising performance in surgical visual question answering (VQA). However, existing surgical VQA data

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

SurGE: A Benchmark and Evaluation Framework for Scientific Survey Generation

DGX agent

arXiv:2508.15658v5 Announce Type: replace Abstract: The rapid growth of academic literature makes the manual creation of scientific surveys increasingly infeasible. While large language models show pr

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

SURGE: SuperBatch Unified Resource-efficient GPU Encoding for Heterogeneous Partitioned Data

DGX agent

arXiv:2605.01060v1 Announce Type: cross Abstract: We present SURGE, a streaming GPU encoding system deployed in production to generate embeddings for over 800 million texts across 40,000 logical parti

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

SwiftChannel: Algorithm-Hardware Co-Design for Deep Learning-Based 5G Channel Estimation

DGX agent

arXiv:2605.01931v1 Announce Type: cross Abstract: Channel estimation is crucial in 5G communication networks for optimizing transmission parameters and ensuring reliable, high-speed communication. How

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Task-Driven Subspace Decomposition for Knowledge Sharing and Isolation in LoRA-based Continual Learning

DGX agent

arXiv:2603.00191v2 Announce Type: replace-cross Abstract: Continual Learning (CL) requires models to sequentially adapt to new tasks without forgetting old knowledge. Recently, Low-Rank Adaptation (Lo

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

TCDA: Thread-Constrained Discourse-Aware Modeling for Conversational Sentiment Quadruple Analysis

DGX agent

arXiv:2605.01717v1 Announce Type: new Abstract: Conversational Aspect-based Sentiment Quadruple Analysis (DiaASQ) needs to capture the complex interrelationships in multiple rounds of dialogues. Exist

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Teaching LLMs Brazilian Healthcare: Injecting Knowledge from Official Clinical Guidelines

DGX agent

arXiv:2605.01077v1 Announce Type: new Abstract: Brazil's Unified Health System (SUS) relies on official clinical guidelines that define diagnostic criteria, treatments, dosages, and monitoring procedu

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

TF1-EN-3M: Three Million Synthetic Moral Fables for Training Small, Open Language Models

DGX agent

arXiv:2504.20605v2 Announce Type: replace Abstract: Moral stories are a time-tested vehicle for transmitting values, yet modern NLP lacks a large, structured corpus that couples coherent narratives wi

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

The 2026 ACII Dyadic Conversations (DaiKon) Workshop & Challenge

DGX agent

arXiv:2605.02672v1 Announce Type: cross Abstract: The 2026 ACII Dyadic Conversations (ACII-DaiKon) Workshop & Challenge introduces a benchmark for modeling interpersonal affect and social dynamics in

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

The Compliance Gap: Why AI Systems Promise to Follow Process Instructions but Don't

DGX agent

arXiv:2605.01771v1 Announce Type: new Abstract: An auditor instructs an AI assistant: 'open each file individually using the Read tool -- no scripts, no agents.' The AI replies 'Yes' -- then issues a

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

The Compliance Trap: How Structural Constraints Degrade Frontier AI Metacognition Under Adversarial Pressure

DGX agent

arXiv:2605.02398v1 Announce Type: cross Abstract: As frontier AI models are deployed in high-stakes decision pipelines, their ability to maintain metacognitive stability -- knowing what they do not kn

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

The Good, the Bad, and the Sampled: a No-Regret Approach to Safe Online Classification

DGX agent

arXiv:2510.01020v2 Announce Type: replace Abstract: We study sequential testing for a binary disease outcome when risk follows an unknown logistic model. At each round, the decision maker may either p

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

The Topology of Multimodal Fusion: Why Current Architectures Fail at Creative Cognition

DGX agent

arXiv:2604.04465v2 Announce Type: replace Abstract: This paper identifies a structural limitation in current multimodal AI architectures that is topological rather than parametric. Contrastive alignme

model-releasesarxiv-cs-ai
5 May 2026
Model Releases

Think2SQL: Reinforce LLM Reasoning Capabilities for Text2SQL

DGX agent

arXiv:2504.15077v5 Announce Type: replace Abstract: Large Language Models (LLMs) can translate natural language into SQL, but small models struggle with multi-table and complex queries in Zero-Shot Le

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

TIME: Temporally Intelligent Meta-reasoning Engine for Context-Triggered Explicit Reasoning

DGX agent

arXiv:2601.05300v2 Announce Type: replace-cross Abstract: Reasoning-oriented language models typically expose explicit reasoning as a long, front-loaded chain of 'thinking' tokens before the main outp

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

TOC-SR: Task-Optimal Compact diffusion for Image Super Resolution

DGX agent

arXiv:2605.02767v1 Announce Type: new Abstract: Diffusion models have recently demonstrated strong performance for image restoration tasks, including super-resolution. However, their large model size

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

TokenTiming: A Dynamic Alignment Method for Universal Speculative Decoding Model Pairs

DGX agent

arXiv:2510.15545v4 Announce Type: replace Abstract: Accelerating the inference of large language models (LLMs) has been a critical challenge in generative AI. Speculative decoding (SD) substantially i

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Toward Culturally Grounded Natural Language Processing

DGX agent

arXiv:2603.26013v2 Announce Type: replace Abstract: Multilingual NLP is often treated as a route to global inclusion, but linguistic coverage and cultural competence frequently diverge. This paper syn

model-releasesarxiv-cs-cl
5 May 2026
← Previous
1…281282283284285…361
Next →