AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
All
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,520 results
Model Releases

Results on the remainder of our benchmarks are in for Kimi K2.6, the new #1 open-weight model on the Vals Index (#8 overall) Kimi K2.6 is co…

DGX agent

Results on the remainder of our benchmarks are in for Kimi K2.6, the new #1 open-weight model on the Vals Index (#8 overall) Kimi K2.6 is competitive with many closed-weight models, at a fraction of t

model-releaseskimi-moonshot--x
28 Apr 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Rethinking Parameter Sharing for LLM Fine-Tuning with Multiple LoRAs

DGX agent

arXiv:2509.25414v2 Announce Type: replace-cross Abstract: Large language models are often adapted using parameter-efficient techniques such as Low-Rank Adaptation (LoRA), formulated as y = W_0x + BAx,

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

ReVSI: Rebuilding Visual Spatial Intelligence Evaluation for Accurate Assessment of VLM 3D Reasoning

DGX agent

arXiv:2604.24300v1 Announce Type: new Abstract: Current evaluations of spatial intelligence can be systematically invalid under modern vision-language model (VLM) settings. First, many benchmarks deri

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

Robust Audio-Text Retrieval via Cross-Modal Attention and Hybrid Loss

DGX agent

arXiv:2604.23323v1 Announce Type: new Abstract: Audio-text retrieval enables semantic alignment between audio content and natural language queries, supporting applications in multimedia search, access

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

RouteNLP: Closed-Loop LLM Routing with Conformal Cascading and Distillation Co-Optimization

DGX agent

arXiv:2604.23577v1 Announce Type: new Abstract: Serving diverse NLP workloads with large language models is costly: at one enterprise partner, inference costs exceeded $200K/month despite over 70% of

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Scalable Agentic Reasoning for Designing Biologics Targeting Intrinsically Disordered Proteins

DGX agent

arXiv:2512.15930v2 Announce Type: replace-cross Abstract: Intrinsically disordered proteins (IDPs) represent crucial therapeutic targets due to their significant role in disease -- approximately 80% o

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Scaling Multi-Node Mixture-of-Experts Inference Using Expert Activation Patterns

DGX agent

arXiv:2604.23150v1 Announce Type: cross Abstract: Most recent state-of-the-art (SOTA) large language models (LLMs) use Mixture-of-Experts (MoE) architectures to scale model capacity without proportion

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Scaling Properties of Continuous Diffusion Spoken Language Models

DGX agent

arXiv:2604.24416v1 Announce Type: cross Abstract: Speech-only spoken language models (SLMs) lag behind text and text-speech models in performance, with recent discrete autoregressive (AR) SLMs indicat

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Scheming Ability in LLM-to-LLM Strategic Interactions

DGX agent

arXiv:2510.12826v2 Announce Type: replace-cross Abstract: As large language model (LLM) agents are deployed autonomously in diverse contexts, evaluating their capacity for strategic deception becomes

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

ScoringBench: A Benchmark for Evaluating Tabular Foundation Models with Proper Scoring Rules

DGX agent

arXiv:2603.29928v2 Announce Type: replace Abstract: Tabular foundation models such as TabPFN and TabICL already produce full predictive distributions, yet prevailing regression benchmarks evaluate the

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Secure On-Premise Deployment of Open-Weights Large Language Models in Radiology: An Isolation-First Architecture with Prospective Pilot Evaluation

DGX agent

arXiv:2604.22768v1 Announce Type: cross Abstract: Purpose: To design, implement, evaluate, and report on the regulatory requirements of a self-hosted LLM infrastructure for radiology adhering to the p

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

See Further, Think Deeper: Advancing VLM's Reasoning Ability with Low-level Visual Cues and Reflection

DGX agent

arXiv:2604.24339v1 Announce Type: cross Abstract: Recent advances in Vision-Language Models (VLMs) have benefited from Reinforcement Learning (RL) for enhanced reasoning. However, existing methods sti

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Seeing Is No Longer Believing: Frontier Image Generation Models, Synthetic Visual Evidence, and Real-World Risk

DGX agent

arXiv:2604.24197v1 Announce Type: cross Abstract: Frontier image generation has moved from artistic synthesis toward synthetic visual evidence. Systems such as GPT Image 2, Nano Banana Pro, Nano Banan

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Semantic Segmentation for Histopathology using Learned Regularization based on Global Proportions

DGX agent

arXiv:2604.24347v1 Announce Type: cross Abstract: In pathology, the spatial distribution and proportions of tissue types are key indicators of disease progression, and are more readily available than

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

SemiGDA: Generative Dual-distribution Alignment for Semi-Supervised Medical Image Segmentation

DGX agent

arXiv:2604.23274v1 Announce Type: new Abstract: Semi-supervised learning addresses label scarcity and high annotation costs in medical image segmentation by exploiting the latent information in unlabe

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

ServImage: An Image Generation and Editing Benchmark from Real-world Commercial Imaging Services

DGX agent

arXiv:2604.24023v1 Announce Type: new Abstract: Recent image generation and editing models demonstrate robust adherence to instructions and high visual quality on academic benchmarks. However, their p

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

SEVerA: Verified Synthesis of Self-Evolving Agents

DGX agent

arXiv:2603.25111v2 Announce Type: replace Abstract: Recent advances have shown the effectiveness of self-evolving LLM agents on tasks such as program repair and scientific discovery. In this paradigm,

model-releasesarxiv-cs-lg
28 Apr 2026
Model Releases

SFT-then-RL Outperforms Mixed-Policy Methods for LLM Reasoning

DGX agent

arXiv:2604.23747v1 Announce Type: cross Abstract: Recent mixed-policy optimization methods for LLM reasoning that interleave or blend supervised and reinforcement learning signals report improvements

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Shape: A Self-Supervised 3D Geometry Foundation Model for Industrial CAD Analysis

DGX agent

arXiv:2604.22826v1 Announce Type: new Abstract: Industrial CAD workflows require robust, generalizable 3D geometric representations supporting accuracy and explainability. We introduce Shape, a self-s

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

ShredBench: Evaluating the Semantic Reasoning Capabilities of Multimodal LLMs in Document Reconstruction

DGX agent

arXiv:2604.23813v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable performance in Visually Rich Document Understanding (VRDU) tasks, but their capabili

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

SIV-Bench: A Video Benchmark for Social Interaction Understanding and Reasoning

DGX agent

arXiv:2506.05425v2 Announce Type: replace-cross Abstract: Understanding social interaction, which encompasses perceiving numerous and subtle multimodal cues, inferring unobservable mental states and r

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

SketchVLM: Vision language models can annotate images to explain thoughts and guide users

DGX agent

arXiv:2604.22875v1 Announce Type: cross Abstract: When answering questions about images, humans naturally point, label, and draw to explain their reasoning. In contrast, modern vision-language models

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Skill Retrieval Augmentation for Agentic AI

DGX agent

arXiv:2604.24594v1 Announce Type: cross Abstract: As large language models (LLMs) evolve into agentic problem solvers, they increasingly rely on external, reusable skills to handle tasks beyond their

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

SMSI: System Model Security Inference: Automated Threat Modeling for Cyber-Physical Systems

DGX agent

arXiv:2604.23905v1 Announce Type: cross Abstract: Threat modeling for cyber-physical systems (CPS) remains a largely manual exercise. This project presents SMSI (System Model Security Inference), a hy

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Snap launches AI Sponsored Snaps, a conversational ad format in Snapchat's Chat tab that lets users talk to brand-specific AI agents for product recommendations (Aisha Malik/TechCrunch)

DGX agent

Aisha Malik / TechCrunch: Snap launches AI Sponsored Snaps, a conversational ad format in Snapchat's Chat tab that lets users talk to brand-specific AI agents for product recommendations — Snapchat an

model-releasestechmeme
28 Apr 2026
Model Releases

SoccerRef-Agents: Multi-Agent System for Automated Soccer Refereeing

DGX agent

arXiv:2604.23392v1 Announce Type: new Abstract: Refereeing is vital in sports, where fair, accurate, and explainable decisions are fundamental. While intelligent assistant technologies are being widel

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

SolarFCD: A Large-Scale Dataset and Benchmark for Solar Fault Classification in Photovoltaic Systems

DGX agent

arXiv:2604.23662v1 Announce Type: new Abstract: The increasing global deployment of solar photovoltaic (PV) systems needs robust, scalable, and automated inspection technologies capable of detecting a

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

SPAGS: Sparse-View Articulated Object Reconstruction from Single State via Planar Gaussian Splatting

DGX agent

arXiv:2511.17092v4 Announce Type: replace Abstract: Articulated objects are ubiquitous in daily environments, and their 3D reconstruction holds great significance across various fields. However, exist

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

Spatiotemporal Degradation-Aware 3D Gaussian Splatting for Realistic Underwater Scene Reconstruction

DGX agent

arXiv:2604.23551v1 Announce Type: new Abstract: Reconstructing realistic underwater scenes from underwater video remains a meaningful yet challenging task in the multimedia domain. The inherent spatio

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

SpecRLBench: A Benchmark for Generalization in Specification-Guided Reinforcement Learning

DGX agent

arXiv:2604.24729v1 Announce Type: new Abstract: Specification-guided reinforcement learning (RL) provides a principled framework for encoding complex, temporally extended tasks using formal specificat

model-releasesarxiv-cs-lg
28 Apr 2026
Model Releases

Speech Enhancement Based on Drifting Models

DGX agent

arXiv:2604.24199v1 Announce Type: cross Abstract: We propose Speech Enhancement based on Drifting Models (DriftSE), a novel generative framework that formulates denoising as an equilibrium problem. Ra

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Speech-FT: Merging Pre-trained And Fine-Tuned Speech Representation Models For Cross-Task Generalization

DGX agent

arXiv:2502.12672v4 Announce Type: replace-cross Abstract: Fine-tuning speech representation models can enhance performance on specific tasks but often compromises their cross-task generalization abili

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Sphere-Depth: A Benchmark for Depth Estimation Methods with Varying Spherical Camera Orientations

DGX agent

arXiv:2604.23432v1 Announce Type: cross Abstract: Reliable depth estimation from spherical images is crucial for 360{eg} vision in robotic navigation and immersive scene understanding. However, the on

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

SSR-Zero: Simple Self-Rewarding Reinforcement Learning for Machine Translation

DGX agent

arXiv:2505.16637v4 Announce Type: replace-cross Abstract: Large language models (LLMs) have recently demonstrated remarkable capabilities in machine translation (MT). However, most advanced MT-specifi

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Stabilizing Efficient Reasoning with Step-Level Advantage Selection

DGX agent

arXiv:2604.24003v1 Announce Type: new Abstract: Large language models (LLMs) achieve strong reasoning performance by allocating substantial computation at inference time, often generating long and ver

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

StoryTR: Narrative-Centric Video Temporal Retrieval with Theory of Mind Reasoning

DGX agent

arXiv:2604.23198v1 Announce Type: new Abstract: Current video moment retrieval excels at action-centric tasks but struggles with narrative content. Models can see extit{what is happening} but fail to

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Strategic Bidding in 6G Spectrum Auctions with Large Language Models

DGX agent

arXiv:2604.24156v1 Announce Type: cross Abstract: Efficient and fair spectrum allocation is a central challenge in 6G networks, where massive connectivity and heterogeneous services continuously compe

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

StratRAG: A Multi-Hop Retrieval Evaluation Dataset for Retrieval-Augmented Generation Systems

DGX agent

arXiv:2604.22757v1 Announce Type: cross Abstract: We introduce StratRAG, an open-source retrieval evaluation dataset for benchmarking Retrieval-Augmented Generation (RAG) systems on multi-hop reasonin

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Stress-Testing Emotional Support Models: Moving from Homogeneous to Diverse Help Seekers

DGX agent

arXiv:2601.07698v2 Announce Type: replace Abstract: As emotional support chatbots have recently gained significant traction across both research and industry, a common evaluation strategy has emerged:

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval

DGX agent

arXiv:2601.20597v2 Announce Type: replace Abstract: Continual Text-to-Video Retrieval (CTVR) is a challenging multimodal continual learning setting, where models must incrementally learn new semantic

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

Structural Pruning of Large Vision Language Models: A Comprehensive Study on Pruning Dynamics, Recovery, and Data Efficiency

DGX agent

arXiv:2604.24380v1 Announce Type: new Abstract: While Large Vision Language Models (LVLMs) demonstrate impressive capabilities, their substantial computational and memory requirements pose deployment

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Supernodes and Halos: Loss-Critical Hubs in LLM Feed-Forward Layers

DGX agent

arXiv:2604.23475v1 Announce Type: cross Abstract: We study the organization of channel-level importance in transformer feed-forward networks (FFNs). Using a Fisher-style loss proxy (LP) based on activ

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

SWE-QA: Can Language Models Answer Repository-level Code Questions?

DGX agent

arXiv:2509.14635v2 Announce Type: replace Abstract: Understanding and reasoning about entire software repositories is an essential capability for intelligent software engineering tools. While existing

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

SwissGov-RSD: A Human-annotated, Cross-lingual Benchmark for Token-level Recognition of Semantic Differences Between Related Documents

DGX agent

arXiv:2512.07538v3 Announce Type: replace Abstract: Recognizing semantic differences across documents is crucial for text generation evaluation and content alignment, especially in cross-lingual setti

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Switch Attention: Towards Dynamic and Fine-grained Hybrid Transformers

DGX agent

arXiv:2603.26380v2 Announce Type: replace Abstract: The attention mechanism has been the core component in modern transformer architectures. However, the computation of standard full attention scales

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

SycoPhantasy: Quantifying Sycophancy and Hallucination in Small Open Weight VLMs for Vision-Language Scoring of Fantasy Characters

DGX agent

arXiv:2604.24346v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly deployed as evaluators in tasks requiring nuanced image understanding, yet their reliability in scoring

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Symmetric Equilibrium Propagation for Thermodynamic Diffusion Training

DGX agent

arXiv:2604.23806v1 Announce Type: cross Abstract: The reverse process in score-based diffusion models is formally equivalent to overdamped Langevin dynamics in a time-dependent energy landscape. In ou

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

SynthPert: Enhancing LLM Biological Reasoning via Synthetic Reasoning Traces for Cellular Perturbation Prediction

DGX agent

arXiv:2509.25346v2 Announce Type: replace Abstract: Predicting cellular responses to genetic perturbations represents a fundamental challenge in systems biology, critical for advancing therapeutic dis

model-releasesarxiv-cs-ai
28 Apr 2026
← Previous
1…386387388389390…470
Next →