AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,520 results
1 May 2026

Math Education Digital Shadows for facilitating learning with LLMs: Math performance, anxiety and confidence in simulated students and AIs

Model ReleasesDGX agent

arXiv:2604.27618v1 Announce Type: new Abstract: To enhance LLMs' impact on math education, we need data on their mathematical prowess and biases across prompts. To fill this gap, we introduce MEDS (Ma

MCPHunt: An Evaluation Framework for Cross-Boundary Data Propagation in Multi-Server MCP Agents

Model ReleasesDGX agent

arXiv:2604.27819v1 Announce Type: new Abstract: Multi-server MCP agents create an information-flow control problem: faithful tool composition can turn individually benign read/write permissions into c

Measurement Risk in Supervised Financial NLP: Rubric and Metric Sensitivity on JF-ICR


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

arXiv:2604.27374v1 Announce Type: new Abstract: As LLMs become credible readers of earnings calls, investor-relations Q&A, guidance, and disclosure language, supervised financial NLP benchmarks increa

MiniCPM-o 4.5: Towards Real-Time Full-Duplex Omni-Modal Interaction

Model ReleasesDGX agent

arXiv:2604.27393v1 Announce Type: new Abstract: Recent progress in multimodal large language models (MLLMs) has brought AI capabilities from static offline data processing to real-time streaming inter

Models Recall What They Violate: Constraint Adherence in Multi-Turn LLM Ideation

Model ReleasesDGX agent

arXiv:2604.28031v1 Announce Type: new Abstract: When researchers iteratively refine ideas with large language models, do the models preserve fidelity to the original objective? We introduce DriftBench

Monitoring Neural Training with Topology: A Footprint-Predictable Collapse Index

Model ReleasesDGX agent

arXiv:2604.26984v1 Announce Type: new Abstract: Representational collapse, where embeddings become anisotropic and lose multi-scale structure, can erode downstream performance long before performance

MultiBLiMP 1.0: A Massively Multilingual Benchmark of Linguistic Minimal Pairs

Model ReleasesDGX agent

arXiv:2504.02768v4 Announce Type: replace Abstract: We introduce MultiBLiMP 1.0, a massively multilingual benchmark of linguistic minimal pairs, covering 101 languages and 2 types of subject-verb agre

NanoKnow: How to Know What Your Language Model Knows

Model ReleasesDGX agent

arXiv:2602.20122v2 Announce Type: replace-cross Abstract: How do large language models (LLMs) know what they know? Answering this question has been difficult because pre-training data is often a 'blac

NashPG: A Policy Gradient Method with Iteratively Refined Regularization for Finding Nash Equilibria

Model ReleasesDGX agent

arXiv:2510.18183v2 Announce Type: replace Abstract: Finding Nash equilibria in two-player zero-sum imperfect-information games remains a central challenge in multi-agent reinforcement learning. Recent

NeocorRAG: Less Irrelevant Information, More Explicit Evidence, and More Effective Recall via Evidence Chains

Model ReleasesDGX agent

arXiv:2604.27852v1 Announce Type: cross Abstract: Although precise recall is a core objective in Retrieval-Augmented Generation (RAG), a critical oversight persists in the field: improvements in retri

ObjectGraph: From Document Injection to Knowledge Traversal -- A Native File Format for the Agentic Era

Model ReleasesDGX agent

arXiv:2604.27820v1 Announce Type: new Abstract: Every document format in existence was designed for a human reader moving linearly through text. Autonomous LLM agents do not read - they retrieve. This

📢 Official Announcement: Qwen Partners with Fireworks AI to Accelerate Access to Qwen Family Models We are pleased to announce a strategic …

Model ReleasesDGX agent

📢 Official Announcement: Qwen Partners with Fireworks AI to Accelerate Access to Qwen Family Models We are pleased to announce a strategic partnership between Qwen and Fireworks AI to deliver optimize

On 5/5 @realDanFu and team will discuss DSV4’s hybrid attention and KV cache efficiency, should be a great session!

Model ReleasesDGX agent

On 5/5 @realDanFu and team will discuss DSV4’s hybrid attention and KV cache efficiency, should be a great session! Join us Tue 5/5: #DeepSeek-V4's hybrid attention + sparse MoE reduces KV cache up to

Once again IP thieves accusing other IP thieves of IP thievery! The nerve!

Model ReleasesDGX agent

Once again IP thieves accusing other IP thieves of IP thievery! The nerve! @ChrisRMcGuire Anthropic claimed OpenAI violated their terms of service by using their API (I suspect they would call it “dis

One week since the launch of GPT-5.5, and it’s already our strongest model launch yet. API revenue is growing more than 2x faster than any p…

Model ReleasesDGX agent

One week since the launch of GPT-5.5, and it’s already our strongest model launch yet. API revenue is growing more than 2x faster than any prior release, while Codex doubled revenue in under seven day

OpenClassGen: A Large-Scale Corpus of Real-World Python Classes for LLM Research

Model ReleasesDGX agent

arXiv:2504.15564v3 Announce Type: replace-cross Abstract: Existing class-level code generation datasets are either synthetic (ClassEval: 100 classes) or insufficient in scale for modern training needs

OR-VSKC: Resolving Visual-Semantic Knowledge Conflicts in Operating Rooms with Synthetic Data-Guided Alignment

Model ReleasesDGX agent

arXiv:2506.22500v2 Announce Type: replace-cross Abstract: Automated identification of surgical safety risks is critical for improving patient outcomes; however, Multimodal Large Language Models (MLLMs

ORFS-agent: Tool-Using Agents for Chip Design Optimization

Model ReleasesDGX agent

arXiv:2506.08332v3 Announce Type: replace Abstract: Machine learning has been widely used to optimize complex engineering workflows across numerous domains. In integrated circuit design, modern flows

Our CEO @jerryjliu0 in @VentureBeat , on what's actually changing in the LLM stack: 'We've really identified that there's a core set of data…

Model ReleasesDGX agent

Our CEO @jerryjliu0 in @VentureBeat , on what's actually changing in the LLM stack: 'We've really identified that there's a core set of data that has been locked up in all these file format containers

Parameter-Efficient Architectural Modifications for Translation-Invariant CNNs

Model ReleasesDGX agent

arXiv:2604.27870v1 Announce Type: new Abstract: Convolutional Neural Networks (CNNs) are widely assumed to be translation-invariant, yet standard architectures exhibit a startling fragility: even a si

People-Centred Medical Image Analysis

Model ReleasesDGX agent

arXiv:2604.26991v1 Announce Type: cross Abstract: Recent advances in data-centric medical AI have produced highly accurate diagnostic systems, but the emphasis on data curation and performance metrics

Perturbation Probing: A Two-Pass-per-Prompt Diagnostic for FFN Behavioral Circuits in Aligned LLMs

Model ReleasesDGX agent

arXiv:2604.27401v1 Announce Type: new Abstract: Perturbation probing generates task-specific causal hypotheses for FFN neurons in large language models using two forward passes per prompt and no backp

PhyCo: Learning Controllable Physical Priors for Generative Motion

Model ReleasesDGX agent

arXiv:2604.28169v1 Announce Type: cross Abstract: Modern video diffusion models excel at appearance synthesis but still struggle with physical consistency: objects drift, collisions lack realistic reb

Physical Foundation Models: Fixed hardware implementations of large-scale neural networks

Model ReleasesDGX agent

arXiv:2604.27911v1 Announce Type: new Abstract: Foundation models are deep neural networks (such as GPT-5, Gemini~3, and Opus~4) trained on large datasets that can perform diverse downstream tasks --

Post-Optimization Adaptive Rank Allocation for LoRA

Model ReleasesDGX agent

arXiv:2604.27796v1 Announce Type: new Abstract: Exponential growth in the scale of modern foundation models has led to the widespread adoption of Low-Rank Adaptation (LoRA) as a parameter-efficient fi

Predicting Upcoming Stuttering Events from Three-Second Audio: Stratified Evaluation Reveals Severity-Selective Precursors, and the Model Deploys Fully On-Device

Model ReleasesDGX agent

arXiv:2604.27279v1 Announce Type: cross Abstract: Audio-based stuttering systems to date have been trained for detection -- what disfluency is present now -- leaving prediction, the capability needed

PRISM: Pre-alignment via Black-box On-policy Distillation for Multimodal Reinforcement Learning

Model ReleasesDGX agent

arXiv:2604.28123v1 Announce Type: cross Abstract: The standard post-training recipe for large multimodal models (LMMs) applies supervised fine-tuning (SFT) on curated demonstrations followed by reinfo

Progressive Multi-Agent Reasoning for Biological Perturbation Prediction

Model ReleasesDGX agent

arXiv:2602.07408v2 Announce Type: replace Abstract: Predicting gene regulation responses to biological perturbations requires reasoning about underlying biological causalities. While large language mo

PVeRA: Probabilistic Vector-Based Random Matrix Adaptation

Model ReleasesDGX agent

arXiv:2512.07703v2 Announce Type: replace Abstract: Large foundation models have emerged in the last years and are pushing performance boundaries for a variety of tasks. Training or even finetuning su

REBENCH: A Procedural, Fair-by-Construction Benchmark for LLMs on Stripped-Binary Types and Names (Extended Version)

Model ReleasesDGX agent

arXiv:2604.27319v1 Announce Type: cross Abstract: Large Language Models (LLMs) have achieved remarkable progress in recent years, driving their adoption across a wide range of domains, including compu

Reduced NEXI protocol for the quantification of human gray matter microstructure on the Connectome 2.0 scanner

Model ReleasesDGX agent

arXiv:2509.09513v2 Announce Type: replace-cross Abstract: Biophysical diffusion MRI models like Neurite Exchange Imaging (NEXI) are essential for probing gray matter microstructure, estimating compart

Reinforced Agent: Inference-Time Feedback for Tool-Calling Agents

Model ReleasesDGX agent

arXiv:2604.27233v1 Announce Type: new Abstract: Tool-calling agents are evaluated on tool selection, parameter accuracy, and scope recognition, yet LLM trajectory assessments remain inherently post-ho

Representative Spectral Correlation Network for Multi-source Remote Sensing Image Classification

Model ReleasesDGX agent

arXiv:2604.27323v1 Announce Type: cross Abstract: Hyperspectral image (HSI) and SAR/LiDAR data offer complementary spectral and structural information for land-cover classification. However, their eff

Researchers detail CopyFail, a now-patched Linux vulnerability that lets unprivileged users gain admin access, as many distributions have yet to add fixes (Dan Goodin/Ars Technica)

Model ReleasesDGX agent

Dan Goodin / Ars Technica: Researchers detail CopyFail, a now-patched Linux vulnerability that lets unprivileged users gain admin access, as many distributions have yet to add fixes — Publicly release

RIHA: Report-Image Hierarchical Alignment for Radiology Report Generation

Model ReleasesDGX agent

arXiv:2604.27559v1 Announce Type: cross Abstract: Radiology report generation (RRG) has emerged as a promising approach to alleviate radiologists' workload and reduce human errors by automatically gen

RL is a bit of a double edged sword: in known territory performance increases, but in unknown territory the model tends to hallucinate that …

Model ReleasesDGX agent

RL is a bit of a double edged sword: in known territory performance increases, but in unknown territory the model tends to hallucinate that it is performing a completely different task it was trained

RoadMapper: A Multi-Agent System for Roadmap Generation of Solving Complex Research Problems

Model ReleasesDGX agent

arXiv:2604.27616v1 Announce Type: new Abstract: People commonly leverage structured content to accelerate knowledge acquisition and research problem solving. Among these, roadmaps guide researchers th

Rocket Report: Falcon Heavy is back; Russia's Soyuz-5 finally debuts

Model ReleasesDGX agent

SpaceX's Falcon Heavy returned to flight this week, marking its first mission of 2026 by launching a heavy-lift payload for the US Space Force. Russia's Soyuz-5 rocket lifted off on April 30 from Baik

RPC-Bench: A Fine-grained Benchmark for Research Paper Comprehension

Model ReleasesDGX agent

arXiv:2601.14289v2 Announce Type: replace-cross Abstract: Understanding research papers remains challenging for foundation models due to specialized scientific discourse and complex figures and tables

RuC: HDL-Agnostic Rule Completion Benchmark Generation

Model ReleasesDGX agent

arXiv:2604.27780v1 Announce Type: cross Abstract: Large Language Models (LLMs) have rapidly improved in performance across code-related tasks, making their integration into Register Transfer Level (RT

SCOPE-FE: Structured Control of Operator and Pairwise Exploration for Feature Engineering

Model ReleasesDGX agent

arXiv:2604.27025v1 Announce Type: cross Abstract: Automatic feature engineering is an effective approach for improving predictive performance in tabular learning. However, expand-and-reduce methods, s

Semantic Variational Bayes Based on Semantic Information G Theory for Solving Latent Variables

Model ReleasesDGX agent

arXiv:2408.13122v2 Announce Type: replace-cross Abstract: The Variational Bayesian method (VB) is used to solve the probability distributions of latent variables with the minimum free energy criterion

Simulating Validity: Modal Decoupling in MLLM Generated Feedback on Science Drawings

Model ReleasesDGX agent

arXiv:2604.26957v1 Announce Type: cross Abstract: In science education, students frequently construct hand-drawn visual models of scientific phenomena. These drawings rely on a visual structure where

Skills-Coach: A Self-Evolving Skill Optimizer via Training-Free GRPO

Model ReleasesDGX agent

arXiv:2604.27488v1 Announce Type: new Abstract: We introduce Skills-Coach, a novel automated framework designed to significantly enhance the self-evolution of skills within Large Language Model (LLM)-

Softmax-GS: Generalized Gaussians Learning When to Blend or Bound

Model ReleasesDGX agent

arXiv:2604.27437v1 Announce Type: new Abstract: 3D Gaussian Splatting (3D GS) is widely adopted for novel view synthesis due to its high training and rendering efficiency. However, its efficiency reli

SpatialGrammar: A Domain-Specific Language for LLM-Based 3D Indoor Scene Generation

Model ReleasesDGX agent

arXiv:2604.27555v1 Announce Type: new Abstract: Automatically generating interactive 3D indoor scenes from natural language is crucial for virtual reality, gaming, and embodied AI. However, existing L

Spectral Dynamic Attention Network for Hyperspectral Image Super-Resolution

Model ReleasesDGX agent

arXiv:2604.27326v1 Announce Type: cross Abstract: Hyperspectral image super-resolution is essential for enhancing the spatial fidelity of HSI data, yet existing deep learning methods often struggle wi

SpecVQA: A Benchmark for Spectral Understanding and Visual Question Answering in Scientific Images

Model ReleasesDGX agent

arXiv:2604.28039v1 Announce Type: new Abstract: Spectra are a prevalent yet highly information-dense form of scientific imagery, presenting substantial challenges to multimodal large language models (

Stable but Wrong: An Inference Limit in Galactic Archaeology

Model ReleasesDGX agent

arXiv:2604.27368v1 Announce Type: new Abstract: Statistical inference in observational science typically relies on a fundamental assumption: as sample size increases and uncertainties decrease, the in

State-Dependent Lyapunov Method for Rank-1 Matrix Factorization

Model ReleasesDGX agent

arXiv:2604.26993v1 Announce Type: cross Abstract: We study gradient descent for rank-1 matrix factorization through a certificate-based viewpoint. The central object is a parameterized quadratic certi

Step-level Optimization for Efficient Computer-use Agents

Model ReleasesDGX agent

arXiv:2604.27151v1 Announce Type: new Abstract: Computer-use agents provide a promising path toward general software automation because they can interact directly with arbitrary graphical user interfa

Targeted Linguistic Analysis of Sign Language Models with Minimal Translation Pairs

Model ReleasesDGX agent

arXiv:2604.27232v1 Announce Type: new Abstract: Models of sign language have historically lagged behind those for spoken language (text and speech). Recent work has greatly improved their performance

Tell-Tale Watermarks for Explanatory Reasoning in Synthetic Media Forensics

Model ReleasesDGX agent

arXiv:2509.05753v2 Announce Type: replace-cross Abstract: The rise of synthetic media has blurred the boundary between reality and fabrication under the evolving power of artificial intelligence, fuel

The Epistemic Planning Domain Definition Language: Official Guideline

Model ReleasesDGX agent

arXiv:2601.20969v3 Announce Type: replace Abstract: Epistemic planning extends (multi-agent) automated planning by making agents' knowledge and beliefs first-class aspects of the planning formalism. O

The Impact of LLM Self-Consistency and Reasoning Effort on Automated Scoring Accuracy and Cost

Model ReleasesDGX agent

arXiv:2604.26954v1 Announce Type: cross Abstract: Strategic model selection and reasoning settings are more effective than ensembling for optimizing automated scoring with large language models (LLMs)

The Inverse-Wisdom Law: Architectural Tribalism and the Consensus Paradox in Agentic Swarms

Model ReleasesDGX agent

arXiv:2604.27274v1 Announce Type: new Abstract: As AI transitions toward multi-agent systems (MAS) to solve complex workflows, research paradigms operate on the axiomatic assumption that agent collabo

The latest crop of models remains below 1% on ARC-AGI-3 -- for now. Where will the scores be by the end of the year?

Model ReleasesDGX agent

The latest crop of models remains below 1% on ARC-AGI-3 -- for now. Where will the scores be by the end of the year? GPT-5.5 & Opus 4.7 on ARC-AGI-3 - GPT-5.5: 0.43% - Opus 4.7: 0.18% We found 3 failu

The new Grok comes in below the latest Chinese open weights models, Grok 4 was at the frontier when released. (& Artificial Analysis: please…

Model ReleasesDGX agent

The new Grok comes in below the latest Chinese open weights models, Grok 4 was at the frontier when released. (& Artificial Analysis: please stop using GDPval-AA which is not a useful test of anything

The TEA Nets framework combines AI and cognitive network science to model targets, events and actors in text

Model ReleasesDGX agent

arXiv:2604.27673v1 Announce Type: new Abstract: We introduce Target-Event-Agent Networks (TEA Nets) as a computational framework to extract subjects (``Agents'), verbs (``Events'), and objects (``Targ

Theory Under Construction: Orchestrating Language Models for Research Software Where the Specification Evolves

Model ReleasesDGX agent

arXiv:2604.27209v1 Announce Type: cross Abstract: Large language models can now generate substantial code and draft research text, but research-software projects require more than either artifact alon

← Previous
1…295296297298299…376
Next →