AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Model Releases

When Prompts Become Payloads: A Framework for Mitigating SQL Injection Attacks in Large Language Model-Driven Applications

DGX agent

arXiv:2605.10176v1 Announce Type: cross Abstract: Natural language interfaces to structured databases are becoming increasingly common, largely due to advances in large language models (LLMs) that ena

model-releasesarxiv-cs-ai
12 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

An Interpretable and Scalable Framework for Evaluating Large Language Models

DGX agent

arXiv:2605.07046v1 Announce Type: cross Abstract: Evaluation of large language models (LLMs) is increasingly critical, yet standard benchmarking methods rely on average accuracy, overlooking both the

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Anatomy of Unlearning: The Dual Impact of Fact Salience and Model Fine-Tuning

DGX agent

arXiv:2602.19612v3 Announce Type: replace Abstract: Machine Unlearning (MU) enables Large Language Models (LLMs) to remove unsafe or outdated information. However, existing work assumes that all facts

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Benchmarking Foundation Models for Renal Lesion Stratification in CT

DGX agent

arXiv:2605.07749v1 Announce Type: new Abstract: The rapid proliferation of open-source medical foundation models (FMs) raises a practical question: how well do their pre-trained representations transf

model-releasesarxiv-cs-cv
11 May 2026
Safety

GPO-V: Jailbreak Diffusion Vision Language Model by Global Probability Optimization

DGX agent

arXiv:2605.07399v1 Announce Type: new Abstract: Diffusion Vision-Language Models (dVLMs), built upon the non-causal foundations of Diffusion Large Language Models (dLLMs), have demonstrated remarkable

safetyarxiv-cs-cv
11 May 2026
Model Releases

Graph Representation Learning Augmented Model Manipulation on Federated Fine-Tuning of LLMs

DGX agent

arXiv:2605.07961v1 Announce Type: new Abstract: Federated fine-tuning (FFT) has emerged as a privacy-preserving paradigm for collaboratively adapting large language models (LLMs). Built upon federated

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Head Similarity: Modeling Structured Whole-Head Appearance Beyond Face Recognition

DGX agent

arXiv:2605.07766v1 Announce Type: new Abstract: Many vision applications require identity consistency beyond strict biometric recognition, especially under non-frontal views or when facial cues are mi

model-releasesarxiv-cs-cv
11 May 2026
Tutorials

How to Train Your Latent Diffusion Language Model Jointly With the Latent Space

DGX agent

arXiv:2605.07933v1 Announce Type: new Abstract: Latent diffusion models offer an attractive alternative to discrete diffusion for non-autoregressive text generation by operating on continuous text rep

tutorialsarxiv-cs-cl
11 May 2026
Safety

Learning Visual Feature-Based World Models via Residual Latent Action

DGX agent

arXiv:2605.07079v1 Announce Type: cross Abstract: World models predict future transitions from observations and actions. Existing works predominantly focus on image generation only. Visual feature-bas

safetyarxiv-cs-ai
11 May 2026
Model Releases

NSMQ Riddles: A Benchmark of Scientific and Mathematical Riddles for Quizzing Large Language Models

DGX agent

arXiv:2605.07051v1 Announce Type: new Abstract: Large Language Models (LLMs) have shown good performance on various science educational benchmarks, demonstrating their potential for use in science and

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Optimizing Language Models for Crosslingual Knowledge Consistency

DGX agent

arXiv:2603.04678v2 Announce Type: replace-cross Abstract: Large language models are known to often exhibit inconsistent knowledge. This is particularly problematic in multilingual scenarios, where mod

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

PerfCoder: Large Language Models for Interpretable Code Performance Optimization

DGX agent

arXiv:2512.14018v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have achieved remarkable progress in automatic code generation, yet their ability to produce high-performance cod

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

RedDiffuser: Auditing Multimodal Safety Failures in Vision-Language Models via Reinforced Diffusion

DGX agent

arXiv:2503.06223v5 Announce Type: replace Abstract: Large Vision-Language Models (VLMs) are increasingly deployed in open-ended environments, where ensuring reliable safety under multimodal inputs is

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

S2S-Arena: Evaluating Paralinguistic Instruction Following in Speech-to-Speech Models

DGX agent

arXiv:2503.05085v2 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have fundamentally reshaped speech-to-speech (S2S) systems, enabling increasingly natural spoken int

model-releasesarxiv-cs-cl
11 May 2026
Agents

Switchcraft: AI Model Router for Agentic Tool Calling

DGX agent

arXiv:2605.07112v1 Announce Type: new Abstract: Agentic AI systems that invoke external tools are powerful but costly, leading developers to default to large models and overspend inference budgets. Mo

agentsarxiv-cs-ai
11 May 2026
Model Releases

Toeplitz MLP Mixers are Low Complexity, Information-Rich Sequence Models

DGX agent

arXiv:2605.06683v1 Announce Type: cross Abstract: Transformer-based large language models are in some respects limited by the quadratic time and space computational complexity of attention. We introdu

model-releasesarxiv-cs-ai
11 May 2026
Research

When Does a Language Model Commit? A Finite-Answer Theory of Pre-Verbalization Commitment

DGX agent

arXiv:2605.06723v1 Announce Type: new Abstract: Language models often generate reasoning before giving a final answer, but the visible answer does not reveal when the model's answer preference became

researcharxiv-cs-ai
11 May 2026
Research

A foundation model of vision, audition, and language for in-silico neuroscience

DGX agent

arXiv:2605.04326v1 Announce Type: cross Abstract: Cognitive neuroscience is fragmented into specialized models, each tailored to specific experimental paradigms, hence preventing a unified model of co

researcharxiv-cs-lg
7 May 2026
Model Releases

A Scalable Multi-Task Model for Virtual Sensors

DGX agent

arXiv:2601.20634v2 Announce Type: replace Abstract: Virtual sensors replace expensive physical sensors in critical applications through machine learning by predicting target signals from available mea

model-releasesarxiv-cs-lg
7 May 2026
Model Releases

Cognitive Twins: Investigating Personalized Thinking Model Building and Its Performance Enhancement with Human-in-the-Loop

DGX agent

arXiv:2605.04761v1 Announce Type: new Abstract: This paper presents the Personalized Thinking Model (PTM), a hierarchical and interpretable learner representation designed for AI supported education.

model-releasesarxiv-cs-lg
7 May 2026
Model Releases

Computer-Aided Design Generation by Cascaded Discrete Diffusion Model

DGX agent

arXiv:2605.05031v1 Announce Type: new Abstract: Recent deep learning approaches seek to automate CAD creation by representing a model as a sequence of discrete commands and parameters, and then genera

model-releasesarxiv-cs-cv
7 May 2026
Model Releases

Constrained Extreme Gradient Boosting for Adapting Reduced-Order Models

DGX agent

arXiv:2605.04130v1 Announce Type: new Abstract: High-fidelity simulations, such as computational fluid dynamics and finite element analysis, are essential for modeling complex engineering systems but

model-releasesarxiv-cs-lg
7 May 2026
Safety

D-OPSD: On-Policy Self-Distillation for Continuously Tuning Step-Distilled Diffusion Models

DGX agent

arXiv:2605.05204v1 Announce Type: new Abstract: The landscape of high-performance image generation models is currently shifting from the inefficient multi-step ones to the efficient few-step counterpa

safetyarxiv-cs-cv
7 May 2026
Safety

Efficiently Aligning Language Models with Online Natural Language Feedback

DGX agent

arXiv:2605.04356v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards has been used to elicit impressive performance from language models in many domains. But, broadly benefic

safetyarxiv-cs-lg
7 May 2026
Research

Ensuring Reliability in Programming Knowledge Tracing: A Re-evaluation of Attention-augmented Models and Experimental Protocols

DGX agent

arXiv:2605.04727v1 Announce Type: new Abstract: Programming Knowledge Tracing (PKT) has recently advanced through hybrid approaches that integrate attention-based feature modeling for code representat

researcharxiv-cs-lg
7 May 2026
Model Releases

Self-Prompting Small Language Models for Privacy-Sensitive Clinical Information Extraction

DGX agent

arXiv:2605.04221v1 Announce Type: new Abstract: Clinical named entity recognition from dental progress notes is challenging because documentation is highly unstructured, domain-specific, and often pri

model-releasesarxiv-cs-cl
7 May 2026
Research

The Impossibility Triangle of Long-Context Modeling

DGX agent

arXiv:2605.05066v1 Announce Type: new Abstract: We identify and prove a fundamental trade-off governing long-sequence models: no model can simultaneously achieve (i) per-step computation independent o

researcharxiv-cs-cl
7 May 2026
Applications

AniMatrix: An Anime Video Generation Model that Thinks in Art, Not Physics

DGX agent

arXiv:2605.03652v1 Announce Type: new Abstract: Video generation models internalize physical realism as their prior. Anime deliberately violates physics: smears, impact frames, chibi shifts; and its t

applicationsarxiv-cs-cv
6 May 2026
Model Releases

Benchmarking Parameter-Efficient Fine-Tuning of Large Language Models for Low-Resource Tajik Text Generation with the Tajik Web Corpus

DGX agent

arXiv:2605.03742v1 Announce Type: new Abstract: This paper is devoted to the adaptation of generative large language models for the Tajik language, a low-resource language with Cyrillic script. To ove

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

Conventional Commit Classification using Large Language Models and Prompt Engineering

DGX agent

arXiv:2605.02033v1 Announce Type: cross Abstract: Conventional commits provide a structured format for writing commit messages, which improves readability, software maintenance, and enables automation

model-releasesarxiv-cs-ai
6 May 2026
Safety

EvoJail: Evolutionary Diverse Jailbreak Prompt Generation for Large Language Models

DGX agent

arXiv:2605.02921v1 Announce Type: cross Abstract: As LLMs continue to shape real-world applications, automated jailbreak generation becomes essential to reveal safety weaknesses and guide model improv

safetyarxiv-cs-lg
6 May 2026
Agents

Foundation-Model-Based Agents in Industrial Automation: Purposes, Capabilities, and Open Challenges

DGX agent

arXiv:2605.02592v1 Announce Type: new Abstract: Foundation models, particularly large language models, are increasingly integrated into agent architectures for industrial tasks such as decision suppor

agentsarxiv-cs-ai
6 May 2026
Model Releases

RFPrompt: Prompt-Based Expert Adaptation of the Large Wireless Model for Modulation Classification

DGX agent

arXiv:2605.03279v1 Announce Type: new Abstract: Automatic modulation classification (AMC) in real-world deployments demands robustness to distribution shifts arising from hardware impairments, unseen

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

StateSMix: Online Lossless Compression via Mamba State Space Models and Sparse N-gram Context Mixing

DGX agent

arXiv:2605.02904v1 Announce Type: new Abstract: We present StateSMix, a fully self-contained lossless compressor that couples an online-trained Mamba-style State Space Model (SSM) with sparse n-gram c

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

Valley3: Scaling Omni Foundation Models for E-commerce

DGX agent

arXiv:2605.01278v1 Announce Type: new Abstract: In this work, we present Valley3, an omni multimodal large language model (MLLM) developed for diverse global e-commerce tasks, with unified understandi

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

VLMaxxing through FrameMogging Training-Free Anti-Recomputation for Video Vision-Language Models

DGX agent

arXiv:2605.03351v1 Announce Type: new Abstract: Video vision-language models (VLMs) keep paying for visual state the stream already told us was stable. The factory wall did not move, but most VLM pipe

model-releasesarxiv-cs-cv
6 May 2026
Model Releases

Component-Aware Self-Speculative Decoding in Hybrid Language Models

DGX agent

arXiv:2605.01106v1 Announce Type: new Abstract: Speculative decoding accelerates autoregressive inference by drafting candidate tokens with a fast model and verifying them in parallel with the target.

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

From Euler to Dormand-Prince: ODE Solvers for Flow Matching Generative Models

DGX agent

arXiv:2605.00836v1 Announce Type: new Abstract: Sampling from Flow Matching generative models requires solving an ordinary differential equation (ODE) whose computational cost is dominated by neural n

model-releasesarxiv-cs-lg
5 May 2026
Applications

Fusing Urban Structure and Semantics: A Conditional Diffusion Model for Cross-City OD Matrix Generation

DGX agent

arXiv:2605.00938v1 Announce Type: new Abstract: Accurate modeling of commuting flows is important for urban governance, traffic planning, and resource allocation. However, the combined influence of in

applicationsarxiv-cs-lg
5 May 2026
Model Releases

Is there 'Secret Sauce'' in Large Language Model Development?

DGX agent

arXiv:2602.07238v2 Announce Type: replace-cross Abstract: Do leading LLM developers possess a proprietary ``secret sauce'', or is LLM performance driven by scaling up compute? Using training and bench

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

jina-vlm: Small Multilingual Vision Language Model

DGX agent

arXiv:2512.04032v3 Announce Type: replace Abstract: We present jina-vlm, a token-efficient 2.4B parameter vision-language model that achieves state-of-the-art multilingual VQA performance among open 2

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Language models recognize dropout and Gaussian noise applied to their activations

DGX agent

arXiv:2604.17465v2 Announce Type: replace Abstract: We provide evidence that language models can detect, localize and, to a certain degree, verbalize the difference between perturbations applied to th

model-releasesarxiv-cs-ai
5 May 2026
Model Releases

Latent Trajectory Dynamics in Large Language Models: A Manifold Evolution Framework with Empirical Validation

DGX agent

arXiv:2505.20340v3 Announce Type: replace Abstract: Understanding how latent representations evolve during generation is a central open problem in large language model interpretability. We introduce e

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

NAKUL-Med: Spectral-Graph State Space Models with Dynamics Kernels for Medical Signals

DGX agent

arXiv:2605.00871v1 Announce Type: cross Abstract: State space models (SSMs) achieve linear-time complexity but struggle with multi-channel physiological signals due to three limitations: fixed kernels

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Object-Level Explanations for Image Geolocation Models: a GeoGuessr use-case

DGX agent

arXiv:2605.00912v1 Announce Type: new Abstract: When humans play geolocation games such as GeoGuessr, they rely on concrete visual cues, such as road markings, vegetation, or architectural details, to

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Video Active Perception: Effective Inference-Time Long-Form Video Understanding with Vision-Language Models

DGX agent

arXiv:2605.01662v1 Announce Type: new Abstract: Large vision-language models (VLMs) have advanced multimodal tasks such as video question answering (QA). However, VLMs face the challenge of selecting

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Visual Implicit Autoregressive Modeling

DGX agent

arXiv:2605.01220v1 Announce Type: new Abstract: Visual Autoregressive Modeling (VAR) based on next-scale prediction achieves strong generation quality, but their explicit deep stacks fix the amount of

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Exploring the System 1 Thinking Capability of Large Reasoning Models

DGX agent

arXiv:2504.10368v4 Announce Type: replace Abstract: This paper explores the system 1 thinking capability of Large Reasoning Models (LRMs), the intuitive ability to respond efficiently with minimal tok

model-releasesarxiv-cs-cl
4 May 2026
← Previous
1…6970717273…1030
Next →