AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,489 results
Model Releases

Cognitive Twins: Investigating Personalized Thinking Model Building and Its Performance Enhancement with Human-in-the-Loop

DGX agent

arXiv:2605.04761v1 Announce Type: new Abstract: This paper presents the Personalized Thinking Model (PTM), a hierarchical and interpretable learner representation designed for AI supported education.

model-releasesarxiv-cs-lg
7 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Computer-Aided Design Generation by Cascaded Discrete Diffusion Model

DGX agent

arXiv:2605.05031v1 Announce Type: new Abstract: Recent deep learning approaches seek to automate CAD creation by representing a model as a sequence of discrete commands and parameters, and then genera

model-releasesarxiv-cs-cv
7 May 2026
Model Releases

Constrained Extreme Gradient Boosting for Adapting Reduced-Order Models

DGX agent

arXiv:2605.04130v1 Announce Type: new Abstract: High-fidelity simulations, such as computational fluid dynamics and finite element analysis, are essential for modeling complex engineering systems but

model-releasesarxiv-cs-lg
7 May 2026
Safety

D-OPSD: On-Policy Self-Distillation for Continuously Tuning Step-Distilled Diffusion Models

DGX agent

arXiv:2605.05204v1 Announce Type: new Abstract: The landscape of high-performance image generation models is currently shifting from the inefficient multi-step ones to the efficient few-step counterpa

safetyarxiv-cs-cv
7 May 2026
Safety

Efficiently Aligning Language Models with Online Natural Language Feedback

DGX agent

arXiv:2605.04356v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards has been used to elicit impressive performance from language models in many domains. But, broadly benefic

safetyarxiv-cs-lg
7 May 2026
Research

Ensuring Reliability in Programming Knowledge Tracing: A Re-evaluation of Attention-augmented Models and Experimental Protocols

DGX agent

arXiv:2605.04727v1 Announce Type: new Abstract: Programming Knowledge Tracing (PKT) has recently advanced through hybrid approaches that integrate attention-based feature modeling for code representat

researcharxiv-cs-lg
7 May 2026
Model Releases

Self-Prompting Small Language Models for Privacy-Sensitive Clinical Information Extraction

DGX agent

arXiv:2605.04221v1 Announce Type: new Abstract: Clinical named entity recognition from dental progress notes is challenging because documentation is highly unstructured, domain-specific, and often pri

model-releasesarxiv-cs-cl
7 May 2026
Research

The Impossibility Triangle of Long-Context Modeling

DGX agent

arXiv:2605.05066v1 Announce Type: new Abstract: We identify and prove a fundamental trade-off governing long-sequence models: no model can simultaneously achieve (i) per-step computation independent o

researcharxiv-cs-cl
7 May 2026
Applications

AniMatrix: An Anime Video Generation Model that Thinks in Art, Not Physics

DGX agent

arXiv:2605.03652v1 Announce Type: new Abstract: Video generation models internalize physical realism as their prior. Anime deliberately violates physics: smears, impact frames, chibi shifts; and its t

applicationsarxiv-cs-cv
6 May 2026
Model Releases

Benchmarking Parameter-Efficient Fine-Tuning of Large Language Models for Low-Resource Tajik Text Generation with the Tajik Web Corpus

DGX agent

arXiv:2605.03742v1 Announce Type: new Abstract: This paper is devoted to the adaptation of generative large language models for the Tajik language, a low-resource language with Cyrillic script. To ove

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

Conventional Commit Classification using Large Language Models and Prompt Engineering

DGX agent

arXiv:2605.02033v1 Announce Type: cross Abstract: Conventional commits provide a structured format for writing commit messages, which improves readability, software maintenance, and enables automation

model-releasesarxiv-cs-ai
6 May 2026
Safety

EvoJail: Evolutionary Diverse Jailbreak Prompt Generation for Large Language Models

DGX agent

arXiv:2605.02921v1 Announce Type: cross Abstract: As LLMs continue to shape real-world applications, automated jailbreak generation becomes essential to reveal safety weaknesses and guide model improv

safetyarxiv-cs-lg
6 May 2026
Industry

Fitting the future: How Breuninger boosted sales with its 'be your own model' AI

DGX agent

“How will this look on me?” It’s the question every online fashion shopper asks, and one that most retailers still can’t answer well. Breuninger, a fashion and lifestyle company based in Germany, thou

industrygoogle-cloud-ai
6 May 2026
Agents

Foundation-Model-Based Agents in Industrial Automation: Purposes, Capabilities, and Open Challenges

DGX agent

arXiv:2605.02592v1 Announce Type: new Abstract: Foundation models, particularly large language models, are increasingly integrated into agent architectures for industrial tasks such as decision suppor

agentsarxiv-cs-ai
6 May 2026
Model Releases

RFPrompt: Prompt-Based Expert Adaptation of the Large Wireless Model for Modulation Classification

DGX agent

arXiv:2605.03279v1 Announce Type: new Abstract: Automatic modulation classification (AMC) in real-world deployments demands robustness to distribution shifts arising from hardware impairments, unseen

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

StateSMix: Online Lossless Compression via Mamba State Space Models and Sparse N-gram Context Mixing

DGX agent

arXiv:2605.02904v1 Announce Type: new Abstract: We present StateSMix, a fully self-contained lossless compressor that couples an online-trained Mamba-style State Space Model (SSM) with sparse n-gram c

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

The last sentence in this abstract is really important, in a way that professional programmers will immediately recognize: the models favore…

DGX agent

The last sentence in this abstract is really important, in a way that professional programmers will immediately recognize: the models favored big single files rather than breaking things into modules.

model-releasesgary-marcus--x
6 May 2026
Model Releases

Valley3: Scaling Omni Foundation Models for E-commerce

DGX agent

arXiv:2605.01278v1 Announce Type: new Abstract: In this work, we present Valley3, an omni multimodal large language model (MLLM) developed for diverse global e-commerce tasks, with unified understandi

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

💫Very happy to release NeuralBench, to benchmark Neuro AI models and datasets in the open! 🧵Thread, 💻Code, 📝White Paper below:

DGX agent

💫Very happy to release NeuralBench, to benchmark Neuro AI models and datasets in the open! 🧵Thread, 💻Code, 📝White Paper below: 🧠 Introducing NeuralBench: a unified, open-source framework to benchmark

model-releasesyann-lecun--x
6 May 2026
Model Releases

VLMaxxing through FrameMogging Training-Free Anti-Recomputation for Video Vision-Language Models

DGX agent

arXiv:2605.03351v1 Announce Type: new Abstract: Video vision-language models (VLMs) keep paying for visual state the stream already told us was stable. The factory wall did not move, but most VLM pipe

model-releasesarxiv-cs-cv
6 May 2026
Model Releases

Component-Aware Self-Speculative Decoding in Hybrid Language Models

DGX agent

arXiv:2605.01106v1 Announce Type: new Abstract: Speculative decoding accelerates autoregressive inference by drafting candidate tokens with a fast model and verifying them in parallel with the target.

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

From Euler to Dormand-Prince: ODE Solvers for Flow Matching Generative Models

DGX agent

arXiv:2605.00836v1 Announce Type: new Abstract: Sampling from Flow Matching generative models requires solving an ordinary differential equation (ODE) whose computational cost is dominated by neural n

model-releasesarxiv-cs-lg
5 May 2026
Applications

Fusing Urban Structure and Semantics: A Conditional Diffusion Model for Cross-City OD Matrix Generation

DGX agent

arXiv:2605.00938v1 Announce Type: new Abstract: Accurate modeling of commuting flows is important for urban governance, traffic planning, and resource allocation. However, the combined influence of in

applicationsarxiv-cs-lg
5 May 2026
Model Releases

Is there 'Secret Sauce'' in Large Language Model Development?

DGX agent

arXiv:2602.07238v2 Announce Type: replace-cross Abstract: Do leading LLM developers possess a proprietary ``secret sauce'', or is LLM performance driven by scaling up compute? Using training and bench

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

jina-vlm: Small Multilingual Vision Language Model

DGX agent

arXiv:2512.04032v3 Announce Type: replace Abstract: We present jina-vlm, a token-efficient 2.4B parameter vision-language model that achieves state-of-the-art multilingual VQA performance among open 2

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Language models recognize dropout and Gaussian noise applied to their activations

DGX agent

arXiv:2604.17465v2 Announce Type: replace Abstract: We provide evidence that language models can detect, localize and, to a certain degree, verbalize the difference between perturbations applied to th

model-releasesarxiv-cs-ai
5 May 2026
Model Releases

Latent Trajectory Dynamics in Large Language Models: A Manifold Evolution Framework with Empirical Validation

DGX agent

arXiv:2505.20340v3 Announce Type: replace Abstract: Understanding how latent representations evolve during generation is a central open problem in large language model interpretability. We introduce e

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

NAKUL-Med: Spectral-Graph State Space Models with Dynamics Kernels for Medical Signals

DGX agent

arXiv:2605.00871v1 Announce Type: cross Abstract: State space models (SSMs) achieve linear-time complexity but struggle with multi-channel physiological signals due to three limitations: fixed kernels

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Object-Level Explanations for Image Geolocation Models: a GeoGuessr use-case

DGX agent

arXiv:2605.00912v1 Announce Type: new Abstract: When humans play geolocation games such as GeoGuessr, they rely on concrete visual cues, such as road markings, vegetation, or architectural details, to

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Video Active Perception: Effective Inference-Time Long-Form Video Understanding with Vision-Language Models

DGX agent

arXiv:2605.01662v1 Announce Type: new Abstract: Large vision-language models (VLMs) have advanced multimodal tasks such as video question answering (QA). However, VLMs face the challenge of selecting

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Visual Implicit Autoregressive Modeling

DGX agent

arXiv:2605.01220v1 Announce Type: new Abstract: Visual Autoregressive Modeling (VAR) based on next-scale prediction achieves strong generation quality, but their explicit deep stacks fix the amount of

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Exploring the System 1 Thinking Capability of Large Reasoning Models

DGX agent

arXiv:2504.10368v4 Announce Type: replace Abstract: This paper explores the system 1 thinking capability of Large Reasoning Models (LRMs), the intuitive ability to respond efficiently with minimal tok

model-releasesarxiv-cs-cl
4 May 2026
Agents

Open sources harnesses powered by open source models

DGX agent

Open sources harnesses powered by open source models I hear this take (usually from the model labs), but that will just mean more people turn to open source models, no? They’re already cheaper, if the

agentsharrison-chase--x
4 May 2026
Model Releases

RSAT: Structured Attribution Makes Small Language Models Faithful Table Reasoners

DGX agent

arXiv:2605.00199v1 Announce Type: new Abstract: When a language model answers a table question, users have no way to verify which cells informed which reasoning steps. We introduce RSAT, a method that

model-releasesarxiv-cs-cl
4 May 2026
Agents

This is probably better messaging than just “own your harness.” Yes, open models will need custom harnesses, but that is a means to an end. …

DGX agent

This is probably better messaging than just “own your harness.” Yes, open models will need custom harnesses, but that is a means to an end. The end is utilizing models without being handcuffed to Anth

agentsharrison-chase--x
4 May 2026
Model Releases

When LLMs Stop Following Steps: A Diagnostic Study of Procedural Execution in Language Models

DGX agent

arXiv:2605.00817v1 Announce Type: new Abstract: Large language models (LLMs) often achieve strong performance on reasoning benchmarks, but final-answer accuracy alone does not show whether they faithf

model-releasesarxiv-cs-cl
4 May 2026
Safety

World Model for Robot Learning: A Comprehensive Survey

DGX agent

arXiv:2605.00080v1 Announce Type: cross Abstract: World models, which are predictive representations of how environments evolve under actions, have become a central component of robot learning. They s

safetyarxiv-cs-cv
4 May 2026
Model Releases

the same model in a different harness can yield much different performance! we've seen this on a few different occasions now - we took gpt-5…

DGX agent

the same model in a different harness can yield much different performance! we've seen this on a few different occasions now - we took gpt-5.2-codex from 52.8% to 66.5% on Terminal-Bench 2.0 (Top 30 t

model-releasesharrison-chase--x
3 May 2026
Research

Decoupling the Benefits of Subword Tokenization for Language Model Training via Byte-level Simulation

DGX agent

arXiv:2604.27263v1 Announce Type: new Abstract: Subword tokenization is an essential part of modern large language models (LLMs), yet its specific contributions to training efficiency and model perfor

researcharxiv-cs-cl
1 May 2026
Model Releases

Do World Action Models Generalize Better than VLAs? A Robustness Study

DGX agent

arXiv:2603.22078v3 Announce Type: replace Abstract: Robot action planning in the real world is challenging as it requires not only understanding the current state of the environment but also predictin

model-releasesarxiv-cs-ro
1 May 2026
Safety

From Prompt to Physical Actuation: Holistic Threat Modeling of LLM-Enabled Robotic Systems

DGX agent

arXiv:2604.27267v1 Announce Type: cross Abstract: As large language models are integrated into autonomous robotic systems for task planning and control, compromised inputs or unsafe model outputs can

safetyarxiv-cs-ai
1 May 2026
Research

Linear Models, Variable Selection, Artificial Intelligence

DGX agent

arXiv:2604.27191v1 Announce Type: cross Abstract: Variable selection in linear regression models has been a problem since hypothesis testing began. Which variables to include or exclude from a model i

researcharxiv-cs-lg
1 May 2026
Model Releases

Mapping the Phase Diagram of the Vicsek Model with Machine Learning

DGX agent

arXiv:2604.28167v1 Announce Type: cross Abstract: In this study, we use machine learning to classify and interpolate the phase structure of the Vicsek flocking model across the three-dimensional param

model-releasesarxiv-cs-lg
1 May 2026
Industry

Model Risk Governance Is Not the Same as Risk Intelligence

DGX agent

Model risk governance and risk intelligence are distinct but complementary practices in machine learning operations. Model risk governance focuses on establishing policies, controls, and compliance fr

industrydatabricks
1 May 2026
Model Releases

NanoKnow: How to Know What Your Language Model Knows

DGX agent

arXiv:2602.20122v2 Announce Type: replace-cross Abstract: How do large language models (LLMs) know what they know? Answering this question has been difficult because pre-training data is often a 'blac

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

📢 Official Announcement: Qwen Partners with Fireworks AI to Accelerate Access to Qwen Family Models We are pleased to announce a strategic …

DGX agent

📢 Official Announcement: Qwen Partners with Fireworks AI to Accelerate Access to Qwen Family Models We are pleased to announce a strategic partnership between Qwen and Fireworks AI to deliver optimize

model-releasesqwen--x
1 May 2026
Safety

Sample-efficient evidence estimation of score based priors for model selection

DGX agent

arXiv:2602.20549v2 Announce Type: replace-cross Abstract: The choice of prior is central to solving ill-posed imaging inverse problems, making it essential to select one consistent with the measuremen

safetyarxiv-cs-cv
1 May 2026
Model Releases

Theory Under Construction: Orchestrating Language Models for Research Software Where the Specification Evolves

DGX agent

arXiv:2604.27209v1 Announce Type: cross Abstract: Large language models can now generate substantial code and draft research text, but research-software projects require more than either artifact alon

model-releasesarxiv-cs-ai
1 May 2026
← Previous
1…8889909192…1261
Next →