AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,405 results
8 May 2026

open models got good enough right as frontier inference pricing started creeping up

AgentsDGX agent

Open-source language models have reached sufficient quality and capability levels just as commercial frontier model providers have begun increasing their inference API pricing. This timing creates a p

7 May 2026

A foundation model of vision, audition, and language for in-silico neuroscience

ResearchDGX agent

arXiv:2605.04326v1 Announce Type: cross Abstract: Cognitive neuroscience is fragmented into specialized models, each tailored to specific experimental paradigms, hence preventing a unified model of co

A Scalable Multi-Task Model for Virtual Sensors

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2601.20634v2 Announce Type: replace Abstract: Virtual sensors replace expensive physical sensors in critical applications through machine learning by predicting target signals from available mea

Cognitive Twins: Investigating Personalized Thinking Model Building and Its Performance Enhancement with Human-in-the-Loop

Model ReleasesDGX agent

arXiv:2605.04761v1 Announce Type: new Abstract: This paper presents the Personalized Thinking Model (PTM), a hierarchical and interpretable learner representation designed for AI supported education.

Computer-Aided Design Generation by Cascaded Discrete Diffusion Model

Model ReleasesDGX agent

arXiv:2605.05031v1 Announce Type: new Abstract: Recent deep learning approaches seek to automate CAD creation by representing a model as a sequence of discrete commands and parameters, and then genera

Constrained Extreme Gradient Boosting for Adapting Reduced-Order Models

Model ReleasesDGX agent

arXiv:2605.04130v1 Announce Type: new Abstract: High-fidelity simulations, such as computational fluid dynamics and finite element analysis, are essential for modeling complex engineering systems but

D-OPSD: On-Policy Self-Distillation for Continuously Tuning Step-Distilled Diffusion Models

SafetyDGX agent

arXiv:2605.05204v1 Announce Type: new Abstract: The landscape of high-performance image generation models is currently shifting from the inefficient multi-step ones to the efficient few-step counterpa

Efficiently Aligning Language Models with Online Natural Language Feedback

SafetyDGX agent

arXiv:2605.04356v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards has been used to elicit impressive performance from language models in many domains. But, broadly benefic

Ensuring Reliability in Programming Knowledge Tracing: A Re-evaluation of Attention-augmented Models and Experimental Protocols

ResearchDGX agent

arXiv:2605.04727v1 Announce Type: new Abstract: Programming Knowledge Tracing (PKT) has recently advanced through hybrid approaches that integrate attention-based feature modeling for code representat

Self-Prompting Small Language Models for Privacy-Sensitive Clinical Information Extraction

Model ReleasesDGX agent

arXiv:2605.04221v1 Announce Type: new Abstract: Clinical named entity recognition from dental progress notes is challenging because documentation is highly unstructured, domain-specific, and often pri

The Impossibility Triangle of Long-Context Modeling

ResearchDGX agent

arXiv:2605.05066v1 Announce Type: new Abstract: We identify and prove a fundamental trade-off governing long-sequence models: no model can simultaneously achieve (i) per-step computation independent o

6 May 2026

AniMatrix: An Anime Video Generation Model that Thinks in Art, Not Physics

ApplicationsDGX agent

arXiv:2605.03652v1 Announce Type: new Abstract: Video generation models internalize physical realism as their prior. Anime deliberately violates physics: smears, impact frames, chibi shifts; and its t

Benchmarking Parameter-Efficient Fine-Tuning of Large Language Models for Low-Resource Tajik Text Generation with the Tajik Web Corpus

Model ReleasesDGX agent

arXiv:2605.03742v1 Announce Type: new Abstract: This paper is devoted to the adaptation of generative large language models for the Tajik language, a low-resource language with Cyrillic script. To ove

Conventional Commit Classification using Large Language Models and Prompt Engineering

Model ReleasesDGX agent

arXiv:2605.02033v1 Announce Type: cross Abstract: Conventional commits provide a structured format for writing commit messages, which improves readability, software maintenance, and enables automation

EvoJail: Evolutionary Diverse Jailbreak Prompt Generation for Large Language Models

SafetyDGX agent

arXiv:2605.02921v1 Announce Type: cross Abstract: As LLMs continue to shape real-world applications, automated jailbreak generation becomes essential to reveal safety weaknesses and guide model improv

Fitting the future: How Breuninger boosted sales with its 'be your own model' AI

IndustryDGX agent

“How will this look on me?” It’s the question every online fashion shopper asks, and one that most retailers still can’t answer well. Breuninger, a fashion and lifestyle company based in Germany, thou

Foundation-Model-Based Agents in Industrial Automation: Purposes, Capabilities, and Open Challenges

AgentsDGX agent

arXiv:2605.02592v1 Announce Type: new Abstract: Foundation models, particularly large language models, are increasingly integrated into agent architectures for industrial tasks such as decision suppor

RFPrompt: Prompt-Based Expert Adaptation of the Large Wireless Model for Modulation Classification

Model ReleasesDGX agent

arXiv:2605.03279v1 Announce Type: new Abstract: Automatic modulation classification (AMC) in real-world deployments demands robustness to distribution shifts arising from hardware impairments, unseen

StateSMix: Online Lossless Compression via Mamba State Space Models and Sparse N-gram Context Mixing

Model ReleasesDGX agent

arXiv:2605.02904v1 Announce Type: new Abstract: We present StateSMix, a fully self-contained lossless compressor that couples an online-trained Mamba-style State Space Model (SSM) with sparse n-gram c

The last sentence in this abstract is really important, in a way that professional programmers will immediately recognize: the models favore…

Model ReleasesDGX agent

The last sentence in this abstract is really important, in a way that professional programmers will immediately recognize: the models favored big single files rather than breaking things into modules.

Valley3: Scaling Omni Foundation Models for E-commerce

Model ReleasesDGX agent

arXiv:2605.01278v1 Announce Type: new Abstract: In this work, we present Valley3, an omni multimodal large language model (MLLM) developed for diverse global e-commerce tasks, with unified understandi

💫Very happy to release NeuralBench, to benchmark Neuro AI models and datasets in the open! 🧵Thread, 💻Code, 📝White Paper below:

Model ReleasesDGX agent

💫Very happy to release NeuralBench, to benchmark Neuro AI models and datasets in the open! 🧵Thread, 💻Code, 📝White Paper below: 🧠 Introducing NeuralBench: a unified, open-source framework to benchmark

VLMaxxing through FrameMogging Training-Free Anti-Recomputation for Video Vision-Language Models

Model ReleasesDGX agent

arXiv:2605.03351v1 Announce Type: new Abstract: Video vision-language models (VLMs) keep paying for visual state the stream already told us was stable. The factory wall did not move, but most VLM pipe

5 May 2026

Component-Aware Self-Speculative Decoding in Hybrid Language Models

Model ReleasesDGX agent

arXiv:2605.01106v1 Announce Type: new Abstract: Speculative decoding accelerates autoregressive inference by drafting candidate tokens with a fast model and verifying them in parallel with the target.

From Euler to Dormand-Prince: ODE Solvers for Flow Matching Generative Models

Model ReleasesDGX agent

arXiv:2605.00836v1 Announce Type: new Abstract: Sampling from Flow Matching generative models requires solving an ordinary differential equation (ODE) whose computational cost is dominated by neural n

Fusing Urban Structure and Semantics: A Conditional Diffusion Model for Cross-City OD Matrix Generation

ApplicationsDGX agent

arXiv:2605.00938v1 Announce Type: new Abstract: Accurate modeling of commuting flows is important for urban governance, traffic planning, and resource allocation. However, the combined influence of in

Is there 'Secret Sauce'' in Large Language Model Development?

Model ReleasesDGX agent

arXiv:2602.07238v2 Announce Type: replace-cross Abstract: Do leading LLM developers possess a proprietary ``secret sauce'', or is LLM performance driven by scaling up compute? Using training and bench

jina-vlm: Small Multilingual Vision Language Model

Model ReleasesDGX agent

arXiv:2512.04032v3 Announce Type: replace Abstract: We present jina-vlm, a token-efficient 2.4B parameter vision-language model that achieves state-of-the-art multilingual VQA performance among open 2

Language models recognize dropout and Gaussian noise applied to their activations

Model ReleasesDGX agent

arXiv:2604.17465v2 Announce Type: replace Abstract: We provide evidence that language models can detect, localize and, to a certain degree, verbalize the difference between perturbations applied to th

Latent Trajectory Dynamics in Large Language Models: A Manifold Evolution Framework with Empirical Validation

Model ReleasesDGX agent

arXiv:2505.20340v3 Announce Type: replace Abstract: Understanding how latent representations evolve during generation is a central open problem in large language model interpretability. We introduce e

NAKUL-Med: Spectral-Graph State Space Models with Dynamics Kernels for Medical Signals

Model ReleasesDGX agent

arXiv:2605.00871v1 Announce Type: cross Abstract: State space models (SSMs) achieve linear-time complexity but struggle with multi-channel physiological signals due to three limitations: fixed kernels

Object-Level Explanations for Image Geolocation Models: a GeoGuessr use-case

Model ReleasesDGX agent

arXiv:2605.00912v1 Announce Type: new Abstract: When humans play geolocation games such as GeoGuessr, they rely on concrete visual cues, such as road markings, vegetation, or architectural details, to

Video Active Perception: Effective Inference-Time Long-Form Video Understanding with Vision-Language Models

Model ReleasesDGX agent

arXiv:2605.01662v1 Announce Type: new Abstract: Large vision-language models (VLMs) have advanced multimodal tasks such as video question answering (QA). However, VLMs face the challenge of selecting

Visual Implicit Autoregressive Modeling

Model ReleasesDGX agent

arXiv:2605.01220v1 Announce Type: new Abstract: Visual Autoregressive Modeling (VAR) based on next-scale prediction achieves strong generation quality, but their explicit deep stacks fix the amount of

4 May 2026

Exploring the System 1 Thinking Capability of Large Reasoning Models

Model ReleasesDGX agent

arXiv:2504.10368v4 Announce Type: replace Abstract: This paper explores the system 1 thinking capability of Large Reasoning Models (LRMs), the intuitive ability to respond efficiently with minimal tok

Open sources harnesses powered by open source models

AgentsDGX agent

Open sources harnesses powered by open source models I hear this take (usually from the model labs), but that will just mean more people turn to open source models, no? They’re already cheaper, if the

RSAT: Structured Attribution Makes Small Language Models Faithful Table Reasoners

Model ReleasesDGX agent

arXiv:2605.00199v1 Announce Type: new Abstract: When a language model answers a table question, users have no way to verify which cells informed which reasoning steps. We introduce RSAT, a method that

This is probably better messaging than just “own your harness.” Yes, open models will need custom harnesses, but that is a means to an end. …

AgentsDGX agent

This is probably better messaging than just “own your harness.” Yes, open models will need custom harnesses, but that is a means to an end. The end is utilizing models without being handcuffed to Anth

When LLMs Stop Following Steps: A Diagnostic Study of Procedural Execution in Language Models

Model ReleasesDGX agent

arXiv:2605.00817v1 Announce Type: new Abstract: Large language models (LLMs) often achieve strong performance on reasoning benchmarks, but final-answer accuracy alone does not show whether they faithf

World Model for Robot Learning: A Comprehensive Survey

SafetyDGX agent

arXiv:2605.00080v1 Announce Type: cross Abstract: World models, which are predictive representations of how environments evolve under actions, have become a central component of robot learning. They s

3 May 2026

the same model in a different harness can yield much different performance! we've seen this on a few different occasions now - we took gpt-5…

Model ReleasesDGX agent

the same model in a different harness can yield much different performance! we've seen this on a few different occasions now - we took gpt-5.2-codex from 52.8% to 66.5% on Terminal-Bench 2.0 (Top 30 t

1 May 2026

Decoupling the Benefits of Subword Tokenization for Language Model Training via Byte-level Simulation

ResearchDGX agent

arXiv:2604.27263v1 Announce Type: new Abstract: Subword tokenization is an essential part of modern large language models (LLMs), yet its specific contributions to training efficiency and model perfor

Do World Action Models Generalize Better than VLAs? A Robustness Study

Model ReleasesDGX agent

arXiv:2603.22078v3 Announce Type: replace Abstract: Robot action planning in the real world is challenging as it requires not only understanding the current state of the environment but also predictin

From Prompt to Physical Actuation: Holistic Threat Modeling of LLM-Enabled Robotic Systems

SafetyDGX agent

arXiv:2604.27267v1 Announce Type: cross Abstract: As large language models are integrated into autonomous robotic systems for task planning and control, compromised inputs or unsafe model outputs can

Linear Models, Variable Selection, Artificial Intelligence

ResearchDGX agent

arXiv:2604.27191v1 Announce Type: cross Abstract: Variable selection in linear regression models has been a problem since hypothesis testing began. Which variables to include or exclude from a model i

Mapping the Phase Diagram of the Vicsek Model with Machine Learning

Model ReleasesDGX agent

arXiv:2604.28167v1 Announce Type: cross Abstract: In this study, we use machine learning to classify and interpolate the phase structure of the Vicsek flocking model across the three-dimensional param

Model Risk Governance Is Not the Same as Risk Intelligence

IndustryDGX agent

Model risk governance and risk intelligence are distinct but complementary practices in machine learning operations. Model risk governance focuses on establishing policies, controls, and compliance fr

NanoKnow: How to Know What Your Language Model Knows

Model ReleasesDGX agent

arXiv:2602.20122v2 Announce Type: replace-cross Abstract: How do large language models (LLMs) know what they know? Answering this question has been difficult because pre-training data is often a 'blac

📢 Official Announcement: Qwen Partners with Fireworks AI to Accelerate Access to Qwen Family Models We are pleased to announce a strategic …

Model ReleasesDGX agent

📢 Official Announcement: Qwen Partners with Fireworks AI to Accelerate Access to Qwen Family Models We are pleased to announce a strategic partnership between Qwen and Fireworks AI to deliver optimize

Sample-efficient evidence estimation of score based priors for model selection

SafetyDGX agent

arXiv:2602.20549v2 Announce Type: replace-cross Abstract: The choice of prior is central to solving ill-posed imaging inverse problems, making it essential to select one consistent with the measuremen

Theory Under Construction: Orchestrating Language Models for Research Software Where the Specification Evolves

Model ReleasesDGX agent

arXiv:2604.27209v1 Announce Type: cross Abstract: Large language models can now generate substantial code and draft research text, but research-software projects require more than either artifact alon

30 Apr 2026

A New Semisupervised Technique for Polarity Analysis using Masked Language Models

ResearchDGX agent

arXiv:2604.26230v1 Announce Type: new Abstract: I developed a new version of Latent Semantic Scaling (LSS) employing word2vec as a masked language model. Unlike original spatial models, it assigns pol

Affective Flow Language Model for Emotional Support Conversation

Model ReleasesDGX agent

arXiv:2602.08826v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have been widely applied to emotional support conversation (ESC). However, complex multi-turn support remains cha

Graph Property Inference in Small Language Models: Effects of Representation and Reasoning Strategy

ResearchDGX agent

arXiv:2603.06635v2 Announce Type: replace Abstract: Recent progress in language modeling has expanded the range of tasks that can be approached through natural language interfaces, including problems

L2RU: a Structured State Space Model with prescribed L2-bound

Model ReleasesDGX agent

arXiv:2503.23818v3 Announce Type: replace-cross Abstract: Structured state-space models (SSMs) have recently emerged as a powerful architecture at the intersection of machine learning and control, fea

ReLoop: Structured Modeling and Behavioral Verification for Reliable LLM-Based Optimization

Model ReleasesDGX agent

arXiv:2602.15983v2 Announce Type: replace-cross Abstract: Large language models (LLMs) can translate natural language into optimization code, but silent failures pose a critical risk: code that execut

The @huggingface CLI now includes a command to find the best open-source models for a given dataset! This is going to be very useful for age…

IndustryDGX agent

Hugging Face introduced a new CLI command that helps users identify the best open-source models suited for their specific datasets, streamlining the model selection process. This feature reduces the m

Towards Redundancy Reduction in Diffusion Models for Efficient Video Super-Resolution

TutorialsDGX agent

arXiv:2509.23980v2 Announce Type: replace Abstract: Diffusion models have recently shown promising results for video super-resolution (VSR). However, directly adapting generative diffusion models to V

29 Apr 2026

Adaptive Meta-Learning Stochastic Gradient Hamiltonian Monte Carlo Simulation for Bayesian Updating of Structural Dynamic Models

ResearchDGX agent

arXiv:2604.25710v1 Announce Type: cross Abstract: In the last few decades, Markov chain Monte Carlo (MCMC) methods have been widely applied to Bayesian updating of structural dynamic models in the fie

BEVal: A Cross-dataset Evaluation Study of BEV Segmentation Models for Autonomous Driving

AgentsDGX agent

arXiv:2408.16322v4 Announce Type: replace Abstract: Current research in semantic bird's-eye view segmentation for autonomous driving focuses solely on optimizing neural network models using a single d

← Previous
1…7071727374…1007
Next →