AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,563 results
Model Releases

Monitoring Neural Training with Topology: A Footprint-Predictable Collapse Index

DGX agent

arXiv:2604.26984v1 Announce Type: new Abstract: Representational collapse, where embeddings become anisotropic and lose multi-scale structure, can erode downstream performance long before performance

model-releasesarxiv-cs-lg
1 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

MultiBLiMP 1.0: A Massively Multilingual Benchmark of Linguistic Minimal Pairs

DGX agent

arXiv:2504.02768v4 Announce Type: replace Abstract: We introduce MultiBLiMP 1.0, a massively multilingual benchmark of linguistic minimal pairs, covering 101 languages and 2 types of subject-verb agre

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

NanoKnow: How to Know What Your Language Model Knows

DGX agent

arXiv:2602.20122v2 Announce Type: replace-cross Abstract: How do large language models (LLMs) know what they know? Answering this question has been difficult because pre-training data is often a 'blac

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

NashPG: A Policy Gradient Method with Iteratively Refined Regularization for Finding Nash Equilibria

DGX agent

arXiv:2510.18183v2 Announce Type: replace Abstract: Finding Nash equilibria in two-player zero-sum imperfect-information games remains a central challenge in multi-agent reinforcement learning. Recent

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

NeocorRAG: Less Irrelevant Information, More Explicit Evidence, and More Effective Recall via Evidence Chains

DGX agent

arXiv:2604.27852v1 Announce Type: cross Abstract: Although precise recall is a core objective in Retrieval-Augmented Generation (RAG), a critical oversight persists in the field: improvements in retri

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

ObjectGraph: From Document Injection to Knowledge Traversal -- A Native File Format for the Agentic Era

DGX agent

arXiv:2604.27820v1 Announce Type: new Abstract: Every document format in existence was designed for a human reader moving linearly through text. Autonomous LLM agents do not read - they retrieve. This

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

📢 Official Announcement: Qwen Partners with Fireworks AI to Accelerate Access to Qwen Family Models We are pleased to announce a strategic …

DGX agent

📢 Official Announcement: Qwen Partners with Fireworks AI to Accelerate Access to Qwen Family Models We are pleased to announce a strategic partnership between Qwen and Fireworks AI to deliver optimize

model-releasesqwen--x
1 May 2026
Model Releases

On 5/5 @realDanFu and team will discuss DSV4’s hybrid attention and KV cache efficiency, should be a great session!

DGX agent

On 5/5 @realDanFu and team will discuss DSV4’s hybrid attention and KV cache efficiency, should be a great session! Join us Tue 5/5: #DeepSeek-V4's hybrid attention + sparse MoE reduces KV cache up to

model-releasestogether-ai--x
1 May 2026
Model Releases

Once again IP thieves accusing other IP thieves of IP thievery! The nerve!

DGX agent

Once again IP thieves accusing other IP thieves of IP thievery! The nerve! @ChrisRMcGuire Anthropic claimed OpenAI violated their terms of service by using their API (I suspect they would call it “dis

model-releasesgary-marcus--x
1 May 2026
Model Releases

One week since the launch of GPT-5.5, and it’s already our strongest model launch yet. API revenue is growing more than 2x faster than any p…

DGX agent

One week since the launch of GPT-5.5, and it’s already our strongest model launch yet. API revenue is growing more than 2x faster than any prior release, while Codex doubled revenue in under seven day

model-releasesopenai--x
1 May 2026
Model Releases

OpenClassGen: A Large-Scale Corpus of Real-World Python Classes for LLM Research

DGX agent

arXiv:2504.15564v3 Announce Type: replace-cross Abstract: Existing class-level code generation datasets are either synthetic (ClassEval: 100 classes) or insufficient in scale for modern training needs

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

OR-VSKC: Resolving Visual-Semantic Knowledge Conflicts in Operating Rooms with Synthetic Data-Guided Alignment

DGX agent

arXiv:2506.22500v2 Announce Type: replace-cross Abstract: Automated identification of surgical safety risks is critical for improving patient outcomes; however, Multimodal Large Language Models (MLLMs

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

ORFS-agent: Tool-Using Agents for Chip Design Optimization

DGX agent

arXiv:2506.08332v3 Announce Type: replace Abstract: Machine learning has been widely used to optimize complex engineering workflows across numerous domains. In integrated circuit design, modern flows

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Our CEO @jerryjliu0 in @VentureBeat , on what's actually changing in the LLM stack: 'We've really identified that there's a core set of data…

DGX agent

Our CEO @jerryjliu0 in @VentureBeat , on what's actually changing in the LLM stack: 'We've really identified that there's a core set of data that has been locked up in all these file format containers

model-releasesjerry-liu--x
1 May 2026
Model Releases

Parameter-Efficient Architectural Modifications for Translation-Invariant CNNs

DGX agent

arXiv:2604.27870v1 Announce Type: new Abstract: Convolutional Neural Networks (CNNs) are widely assumed to be translation-invariant, yet standard architectures exhibit a startling fragility: even a si

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

People-Centred Medical Image Analysis

DGX agent

arXiv:2604.26991v1 Announce Type: cross Abstract: Recent advances in data-centric medical AI have produced highly accurate diagnostic systems, but the emphasis on data curation and performance metrics

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Perturbation Probing: A Two-Pass-per-Prompt Diagnostic for FFN Behavioral Circuits in Aligned LLMs

DGX agent

arXiv:2604.27401v1 Announce Type: new Abstract: Perturbation probing generates task-specific causal hypotheses for FFN neurons in large language models using two forward passes per prompt and no backp

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

PhyCo: Learning Controllable Physical Priors for Generative Motion

DGX agent

arXiv:2604.28169v1 Announce Type: cross Abstract: Modern video diffusion models excel at appearance synthesis but still struggle with physical consistency: objects drift, collisions lack realistic reb

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Physical Foundation Models: Fixed hardware implementations of large-scale neural networks

DGX agent

arXiv:2604.27911v1 Announce Type: new Abstract: Foundation models are deep neural networks (such as GPT-5, Gemini~3, and Opus~4) trained on large datasets that can perform diverse downstream tasks --

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

Post-Optimization Adaptive Rank Allocation for LoRA

DGX agent

arXiv:2604.27796v1 Announce Type: new Abstract: Exponential growth in the scale of modern foundation models has led to the widespread adoption of Low-Rank Adaptation (LoRA) as a parameter-efficient fi

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Predicting Upcoming Stuttering Events from Three-Second Audio: Stratified Evaluation Reveals Severity-Selective Precursors, and the Model Deploys Fully On-Device

DGX agent

arXiv:2604.27279v1 Announce Type: cross Abstract: Audio-based stuttering systems to date have been trained for detection -- what disfluency is present now -- leaving prediction, the capability needed

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

PRISM: Pre-alignment via Black-box On-policy Distillation for Multimodal Reinforcement Learning

DGX agent

arXiv:2604.28123v1 Announce Type: cross Abstract: The standard post-training recipe for large multimodal models (LMMs) applies supervised fine-tuning (SFT) on curated demonstrations followed by reinfo

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Progressive Multi-Agent Reasoning for Biological Perturbation Prediction

DGX agent

arXiv:2602.07408v2 Announce Type: replace Abstract: Predicting gene regulation responses to biological perturbations requires reasoning about underlying biological causalities. While large language mo

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

PVeRA: Probabilistic Vector-Based Random Matrix Adaptation

DGX agent

arXiv:2512.07703v2 Announce Type: replace Abstract: Large foundation models have emerged in the last years and are pushing performance boundaries for a variety of tasks. Training or even finetuning su

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

REBENCH: A Procedural, Fair-by-Construction Benchmark for LLMs on Stripped-Binary Types and Names (Extended Version)

DGX agent

arXiv:2604.27319v1 Announce Type: cross Abstract: Large Language Models (LLMs) have achieved remarkable progress in recent years, driving their adoption across a wide range of domains, including compu

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

Reduced NEXI protocol for the quantification of human gray matter microstructure on the Connectome 2.0 scanner

DGX agent

arXiv:2509.09513v2 Announce Type: replace-cross Abstract: Biophysical diffusion MRI models like Neurite Exchange Imaging (NEXI) are essential for probing gray matter microstructure, estimating compart

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Reinforced Agent: Inference-Time Feedback for Tool-Calling Agents

DGX agent

arXiv:2604.27233v1 Announce Type: new Abstract: Tool-calling agents are evaluated on tool selection, parameter accuracy, and scope recognition, yet LLM trajectory assessments remain inherently post-ho

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Representative Spectral Correlation Network for Multi-source Remote Sensing Image Classification

DGX agent

arXiv:2604.27323v1 Announce Type: cross Abstract: Hyperspectral image (HSI) and SAR/LiDAR data offer complementary spectral and structural information for land-cover classification. However, their eff

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

Researchers detail CopyFail, a now-patched Linux vulnerability that lets unprivileged users gain admin access, as many distributions have yet to add fixes (Dan Goodin/Ars Technica)

DGX agent

Dan Goodin / Ars Technica: Researchers detail CopyFail, a now-patched Linux vulnerability that lets unprivileged users gain admin access, as many distributions have yet to add fixes — Publicly release

model-releasestechmeme
1 May 2026
Model Releases

RIHA: Report-Image Hierarchical Alignment for Radiology Report Generation

DGX agent

arXiv:2604.27559v1 Announce Type: cross Abstract: Radiology report generation (RRG) has emerged as a promising approach to alleviate radiologists' workload and reduce human errors by automatically gen

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

RL is a bit of a double edged sword: in known territory performance increases, but in unknown territory the model tends to hallucinate that …

DGX agent

RL is a bit of a double edged sword: in known territory performance increases, but in unknown territory the model tends to hallucinate that it is performing a completely different task it was trained

model-releasesfrancois-chollet--x
1 May 2026
Model Releases

RoadMapper: A Multi-Agent System for Roadmap Generation of Solving Complex Research Problems

DGX agent

arXiv:2604.27616v1 Announce Type: new Abstract: People commonly leverage structured content to accelerate knowledge acquisition and research problem solving. Among these, roadmaps guide researchers th

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

Rocket Report: Falcon Heavy is back; Russia's Soyuz-5 finally debuts

DGX agent

SpaceX's Falcon Heavy returned to flight this week, marking its first mission of 2026 by launching a heavy-lift payload for the US Space Force. Russia's Soyuz-5 rocket lifted off on April 30 from Baik

model-releasesars-technica
1 May 2026
Model Releases

RPC-Bench: A Fine-grained Benchmark for Research Paper Comprehension

DGX agent

arXiv:2601.14289v2 Announce Type: replace-cross Abstract: Understanding research papers remains challenging for foundation models due to specialized scientific discourse and complex figures and tables

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

RuC: HDL-Agnostic Rule Completion Benchmark Generation

DGX agent

arXiv:2604.27780v1 Announce Type: cross Abstract: Large Language Models (LLMs) have rapidly improved in performance across code-related tasks, making their integration into Register Transfer Level (RT

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

SCOPE-FE: Structured Control of Operator and Pairwise Exploration for Feature Engineering

DGX agent

arXiv:2604.27025v1 Announce Type: cross Abstract: Automatic feature engineering is an effective approach for improving predictive performance in tabular learning. However, expand-and-reduce methods, s

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

Semantic Variational Bayes Based on Semantic Information G Theory for Solving Latent Variables

DGX agent

arXiv:2408.13122v2 Announce Type: replace-cross Abstract: The Variational Bayesian method (VB) is used to solve the probability distributions of latent variables with the minimum free energy criterion

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Simulating Validity: Modal Decoupling in MLLM Generated Feedback on Science Drawings

DGX agent

arXiv:2604.26957v1 Announce Type: cross Abstract: In science education, students frequently construct hand-drawn visual models of scientific phenomena. These drawings rely on a visual structure where

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Skills-Coach: A Self-Evolving Skill Optimizer via Training-Free GRPO

DGX agent

arXiv:2604.27488v1 Announce Type: new Abstract: We introduce Skills-Coach, a novel automated framework designed to significantly enhance the self-evolution of skills within Large Language Model (LLM)-

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

Softmax-GS: Generalized Gaussians Learning When to Blend or Bound

DGX agent

arXiv:2604.27437v1 Announce Type: new Abstract: 3D Gaussian Splatting (3D GS) is widely adopted for novel view synthesis due to its high training and rendering efficiency. However, its efficiency reli

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

SpatialGrammar: A Domain-Specific Language for LLM-Based 3D Indoor Scene Generation

DGX agent

arXiv:2604.27555v1 Announce Type: new Abstract: Automatically generating interactive 3D indoor scenes from natural language is crucial for virtual reality, gaming, and embodied AI. However, existing L

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Spectral Dynamic Attention Network for Hyperspectral Image Super-Resolution

DGX agent

arXiv:2604.27326v1 Announce Type: cross Abstract: Hyperspectral image super-resolution is essential for enhancing the spatial fidelity of HSI data, yet existing deep learning methods often struggle wi

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

SpecVQA: A Benchmark for Spectral Understanding and Visual Question Answering in Scientific Images

DGX agent

arXiv:2604.28039v1 Announce Type: new Abstract: Spectra are a prevalent yet highly information-dense form of scientific imagery, presenting substantial challenges to multimodal large language models (

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Stable but Wrong: An Inference Limit in Galactic Archaeology

DGX agent

arXiv:2604.27368v1 Announce Type: new Abstract: Statistical inference in observational science typically relies on a fundamental assumption: as sample size increases and uncertainties decrease, the in

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

State-Dependent Lyapunov Method for Rank-1 Matrix Factorization

DGX agent

arXiv:2604.26993v1 Announce Type: cross Abstract: We study gradient descent for rank-1 matrix factorization through a certificate-based viewpoint. The central object is a parameterized quadratic certi

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

Step-level Optimization for Efficient Computer-use Agents

DGX agent

arXiv:2604.27151v1 Announce Type: new Abstract: Computer-use agents provide a promising path toward general software automation because they can interact directly with arbitrary graphical user interfa

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Targeted Linguistic Analysis of Sign Language Models with Minimal Translation Pairs

DGX agent

arXiv:2604.27232v1 Announce Type: new Abstract: Models of sign language have historically lagged behind those for spoken language (text and speech). Recent work has greatly improved their performance

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

Tell-Tale Watermarks for Explanatory Reasoning in Synthetic Media Forensics

DGX agent

arXiv:2509.05753v2 Announce Type: replace-cross Abstract: The rise of synthetic media has blurred the boundary between reality and fabrication under the evolving power of artificial intelligence, fuel

model-releasesarxiv-cs-ai
1 May 2026
← Previous
1…370371372373374…471
Next →