AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,975 results
Model Releases

Benchmarking Fine-tuning and Retrieval Strategies for a Multimodal Language Model on the NRC Reactor Operator Licensing Examination

DGX agent

arXiv:2607.22067v1 Announce Type: new Abstract: The integration of large language models (LLMs) into the nuclear power industry requires outputs grounded in domain-specific knowledge. This study evalu

model-releasesarxiv-cs-cl
27 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Correlation-Aware and Gaussianity-Preserving Robust Latent Angular Watermarking for Diffusion Models

DGX agent

arXiv:2607.22386v1 Announce Type: new Abstract: Latent domain watermarking for diffusion models embeds watermarks directly into the latent prior, enjoying non-intrusiveness to model parameters and sea

researcharxiv-cs-cv
27 Jul 2026
Research

InnoText: A Unified Model for Visual Text Generation and Editing

DGX agent

arXiv:2607.22101v1 Announce Type: new Abstract: Diffusion models have recently achieved remarkable success in high-fidelity image synthesis, yet their application to visual text generation and editing

researcharxiv-cs-cv
27 Jul 2026
Research

Progress Reward Modeling for Robotic Learning: A Comprehensive Survey

DGX agent

arXiv:2607.21655v1 Announce Type: cross Abstract: Robotic learning takes place in dynamic environments with large behavior spaces. A terminal success signal only tells the robot whether the task is co

researcharxiv-cs-cl
27 Jul 2026
Model Releases

A Sovereign, Open-Source Foundation Model for German and English

DGX agent

arXiv:2607.09424v3 Announce Type: replace-cross Abstract: We present Soofi S 30B-A3B, a sovereign, open-source Mixture-of-Experts (MoE) hybrid Mamba Transformer foundation model for German and English

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

An LLM-Driven Workflow for Automated Process Control Strategy Generation and Tuning from Dynamic Process Models

DGX agent

arXiv:2607.21292v1 Announce Type: new Abstract: We present a structured large-language-model-driven workflow for automated multi-variable control design from dynamic process models. The workflow decom

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Expectation Alignment of Language Models for Real-World User Expectations

DGX agent

arXiv:2607.20485v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated remarkable performance on standard benchmarks, yet it remains largely unexplored whether they truly meet

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

HalluScope: Fine-grained Hallucination Diagnosis for Multimodal Large Language Models

DGX agent

arXiv:2607.21105v1 Announce Type: new Abstract: Although Multimodal Large Language Models have achieved strong performance across a wide range of vision-language tasks, they still suffer from hallucin

model-releasesarxiv-cs-cv
24 Jul 2026
Model Releases

Incomplete Prompt Jailbreaks in Large Language Models

DGX agent

arXiv:2607.20473v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly released as open-weight models with safeguards against harmful requests. Nevertheless, sentence completion

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Moir: Let the Model Direct Its Own Story for Robust Cross-Domain Knowledge Editing

DGX agent

arXiv:2607.20433v1 Announce Type: cross Abstract: While language models remain frozen at their training state, the world evolves continuously. Knowledge editing has emerged as a key alternative to ful

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Self-Evolving Recommendation System: End-To-End Autonomous Model Optimization With LLM Agents

DGX agent

arXiv:2602.10226v2 Announce Type: replace-cross Abstract: Optimizing large-scale machine learning systems, such as recommendation models for global video platforms, requires navigating a massive hyper

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Statistical Inference for Generative Model Comparison

DGX agent

arXiv:2501.18897v4 Announce Type: replace-cross Abstract: Generative models have achieved remarkable success across a range of applications, yet their evaluation still lacks principled uncertainty qua

model-releasesarxiv-cs-lg
24 Jul 2026
Research

Causal dictionary learning reveals and validates transcription-factor binding features in genomic language models

DGX agent

arXiv:2607.19618v1 Announce Type: cross Abstract: Genomic language models achieve strong performance across regulatory-genomics tasks, yet what these models internally represent remains opaque, and th

researcharxiv-cs-ai
23 Jul 2026
Research

Comparing Model-agnostic Feature Selection Methods through Relative Efficiency

DGX agent

arXiv:2508.14268v2 Announce Type: replace-cross Abstract: Feature selection and importance estimation in a model-agnostic setting is an ongoing challenge of significant interest. Wrapper methods are c

researcharxiv-cs-lg
23 Jul 2026
Model Releases

Good Practice Guide for quantifying uncertainties for machine learning models applied to photoplethysmography signals

DGX agent

arXiv:2607.19999v1 Announce Type: new Abstract: This Good Practice Guide presents work done in the QUMPHY project (Uncertainty quantification for machine learning models applied to photoplethysmograph

model-releasesarxiv-cs-lg
23 Jul 2026
Model Releases

KineBench: Benchmarking Embodied World Models via IDM-Free Kinematic Grounding

DGX agent

arXiv:2607.19876v1 Announce Type: new Abstract: Evaluating the physical consistency of embodied world models(EWMs) is a critical open challenge. While closed-loop evaluation via simulator rollouts off

model-releasesarxiv-cs-ro
23 Jul 2026
Safety

Masked Visual Actions for Unified World Modeling

DGX agent

arXiv:2607.19343v1 Announce Type: new Abstract: Video models absorb rich priors over how the visual world moves, interacts, and responds to contact, making them promising substrates for robotic world

safetyarxiv-cs-cv
23 Jul 2026
Model Releases

Reading and Steering Representations of Materials-Science Mechanisms in an Open-Weight Language Model

DGX agent

arXiv:2607.20058v1 Announce Type: new Abstract: Large language models can answer scientific questions, yet a correct output does not reveal whether the model represents or uses the governing physics.

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

When Visual Evidence is Ambiguous: Pareidolia as a Diagnostic Probe for Vision Models

DGX agent

arXiv:2603.03989v3 Announce Type: replace-cross Abstract: When visual evidence is ambiguous, vision models must decide how to interpret face-like patterns. Face pareidolia, the perception of faces in

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

Xiaomi-Robotics-1: Scaling Vision-Language-Action Models with over 100K Hours of Real-World Trajectories

DGX agent

arXiv:2607.15330v2 Announce Type: replace Abstract: We present Xiaomi-Robotics-1, a foundational vision-language-action (VLA) model capable of (1) following diverse language instructions to perform a

model-releasesarxiv-cs-ro
23 Jul 2026
Model Releases

Advancing Multimodal Judge Models through a Capability-Oriented Benchmark and MCTS-Driven Data Generation

DGX agent

arXiv:2603.00546v2 Announce Type: replace Abstract: Using Multimodal Large Language Models (MLLMs) as judges to achieve precise and consistent evaluations has gradually become an emerging paradigm acr

model-releasesarxiv-cs-ai
16 Jul 2026
Model Releases

Graded Entity-Familiarity Readouts in Language Models: Polish Adaptation, Cross-Language Robustness, and Refusal Steering

DGX agent

arXiv:2607.13568v1 Announce Type: cross Abstract: Can a language model estimate its familiarity with an entity before generating an answer? We study activations at the final prompt token in twelve ins

model-releasesarxiv-cs-lg
16 Jul 2026
Model Releases

MxGPS: Multiplex Graph Transformers for a Power Grid Foundation Model

DGX agent

arXiv:2607.13763v1 Announce Type: cross Abstract: Single-task fine-tuning of graph neural networks (GNNs) for power grid problems exhibits a systematic failure mode: models that achieve the lowest in-

model-releasesarxiv-cs-ai
16 Jul 2026
Model Releases

S-squared-VLA: Decoupling Semantic and Spatial Streams in Vision-Language-Action Models for Autonomous Driving

DGX agent

arXiv:2607.13926v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have demonstrated remarkable potential for high-level reasoning in autonomous driving, yet they fundamentally struggle to

model-releasesarxiv-cs-ro
16 Jul 2026
Research

Self-Improving is Often Sudden: Enlightenment-style Finetuning for Large-Scale Models

DGX agent

arXiv:2607.13395v1 Announce Type: new Abstract: The pursuit of autonomously self-improving models has attracted growing interest in the era of large-scale foundation models. Drawing inspiration from t

researcharxiv-cs-lg
16 Jul 2026
Agents

Amplitude-Only FFN Intervention for Tool-Structured LLM Inference Method: Gated Evaluation Protocol, and Cross-Model Empirical Results

DGX agent

arXiv:2607.11183v2 Announce Type: replace Abstract: Large language models increasingly operate as tool-using agents, where small format, argument, or function-call errors can invalidate otherwise plau

agentsarxiv-cs-cl
15 Jul 2026
Model Releases

Belief-reality separation lives in routing over a shared value slot in language models

DGX agent

arXiv:2607.11945v1 Announce Type: new Abstract: Capable language models hold what a character believes apart from what is true: told 'Anna believes the cup is blue; in reality it is red,' they answer

model-releasesarxiv-cs-cl
15 Jul 2026
Tutorials

Can a Language Model Learn Facts Continually in Its Weights?

DGX agent

arXiv:2607.11020v2 Announce Type: replace Abstract: Continual learning promises a language model that keeps acquiring knowledge after training, with each new fact written into its weights. Whether wei

tutorialsarxiv-cs-cl
15 Jul 2026
Research

Saturation Makes Quantization Error Additive: A Coverage Model with a Certificate

DGX agent

arXiv:2607.12266v1 Announce Type: new Abstract: Mixed-precision quantization must decide which parts of a model to keep at higher precision. A common premise, shared by sensitivity-based methods such

researcharxiv-cs-lg
15 Jul 2026
Model Releases

So Many Opinions, So Many LLMs: Comparing Large Language Models to Traditional Machine Learning for Open- Ended Survey Analysis

DGX agent

arXiv:2607.11890v1 Announce Type: cross Abstract: Open-ended surveys offer valuable insights, but they are notoriously difficult to analyze at scale. Building on previous work that employed traditiona

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

Verifier-Based Reinforcement Fine-Tuning of Reasoning Models for Thermal Energy Storage Control

DGX agent

arXiv:2607.12856v1 Announce Type: new Abstract: Buildings are expected to shift cooling loads in response to grid conditions. Thermal energy storage (TES) enables this shift, but scheduling it well re

model-releasesarxiv-cs-lg
15 Jul 2026
Model Releases

VisCo: Leveraging Large Language Models as Intrinsic Encoders for Visual Token Compression

DGX agent

arXiv:2607.12756v1 Announce Type: new Abstract: Vision-language models (VLMs) process large numbers of visual tokens, resulting in substantial inference latency and memory overhead. This has motivated

model-releasesarxiv-cs-cv
15 Jul 2026
Model Releases

Joint Bayesian Parameter and Model Order Estimation for Low-Rank Probability Mass Tensors

DGX agent

arXiv:2410.06329v4 Announce Type: replace-cross Abstract: Obtaining a reliable estimate of the joint probability mass function (PMF) of a set of random variables from observed data is a significant ob

model-releasesarxiv-cs-lg
10 Jul 2026
Model Releases

InfraQR: Edge-Placed QR-Inspired Structured Patch Attacks on Infrared Vision-Language Models

DGX agent

arXiv:2607.07288v1 Announce Type: new Abstract: Infrared vision-language models are increasingly used for perception under low-light and adverse visual conditions, yet their robustness to localized st

model-releasesarxiv-cs-cv
9 Jul 2026
Model Releases

Multi-Agent Robotic Control with Onboard Vision-Language Models

DGX agent

arXiv:2607.07403v1 Announce Type: cross Abstract: Vision Language Models (VLMs) and Vision Language Action (VLA) models have shown promise in robotic control. Yet, they face significant challenges reg

model-releasesarxiv-cs-ro
9 Jul 2026
Model Releases

Thinking Ahead: Foresight Intelligence in MLLMs and World Model

DGX agent

arXiv:2511.18735v3 Announce Type: replace-cross Abstract: In this work, we define Foresight Intelligence as the capability to anticipate and interpret future events-an ability essential for applicatio

model-releasesarxiv-cs-ai
9 Jul 2026
Applications

FedDAF: Federated Domain Adaptation Using Model Functional Distance

DGX agent

arXiv:2509.11819v2 Announce Type: replace-cross Abstract: Federated Domain Adaptation (FDA) is a federated learning (FL) approach that improves model performance at the target client by collaborating

applicationsarxiv-cs-cv
8 Jul 2026
Safety

Distribution-free Deviation Bounds and The Role of Domain Knowledge in Learning via Model Selection with Cross-validation Risk Estimation

DGX agent

arXiv:2303.08777v3 Announce Type: replace-cross Abstract: Cross-validation is one of the most widely used tools for risk estimation and model selection in statistics and machine learning, yet its theo

safetyarxiv-cs-lg
7 Jul 2026
Model Releases

Don't Blame the Large Language Model: How Scaffolding Evolution Shapes Coding Agent Quality

DGX agent

arXiv:2607.03691v1 Announce Type: cross Abstract: Coding agents, autonomous systems that use large language models (LLMs) to resolve software engineering tasks, rely on agentic scaffolding: a middlewa

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

DynaWM: A Base-VLA-Guided World Foundation Model for Moving-Object Manipulation

DGX agent

arXiv:2607.02604v1 Announce Type: new Abstract: Although vision-language-action (VLA) models have received widespread attention, many challenges remain in manipulating dynamic moving objects. In most

model-releasesarxiv-cs-cv
7 Jul 2026
Safety

ELBO-T2IAlign: A Generic ELBO-Based Method for Calibrating Pixel-level Text-Image Alignment in Diffusion Models

DGX agent

arXiv:2506.09740v2 Announce Type: replace-cross Abstract: Diffusion models excel at image generation. Recent studies have shown that these models not only generate high-quality images but also encode

safetyarxiv-cs-ai
7 Jul 2026
Agents

evalci: A Python Library for Statistically Rigorous Comparison of Language Model Evaluations

DGX agent

arXiv:2607.04429v1 Announce Type: cross Abstract: The dominant practice in language model evaluation is to report a single accuracy number per model and declare the higher one better, without testing

agentsarxiv-cs-ai
7 Jul 2026
Model Releases

Prior Bias in Vision Language Models on UML Diagram Interpretation

DGX agent

arXiv:2607.02853v1 Announce Type: new Abstract: Vision Language Models (VLMs) are increasingly applied to software engineering artifacts, especially UML class diagrams whose meaning depends on visual

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

Punching Above Their Weight: Classification-Head Fine-Tuning of Tiny Language Models (TLMs) for Verifiable Multiple-Choice Tasks

DGX agent

arXiv:2607.03801v1 Announce Type: cross Abstract: We define Tiny Language Models (TLMs) as models below roughly 3B parameters that fit on mainstream consumer devices. We study how to adapt them for an

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Retroactive Chain-of-Thought (RetroCoT): Forensic Reconstruction Prompts as a Safety Diagnostic Across Model Generations

DGX agent

arXiv:2607.04645v1 Announce Type: cross Abstract: Safety alignment in large language models is typically evaluated against direct, imperative harmful requests. We show that this alignment is highly co

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models

DGX agent

arXiv:2607.05365v1 Announce Type: cross Abstract: Streaming speech-to-speech language models aim to answer spoken queries directly with synthetic speech. However, standard speech and text benchmarks d

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

SteelBench: Evaluating Vision-Language Models in Real-World Industrial Environments

DGX agent

arXiv:2607.05264v1 Announce Type: new Abstract: Existing video benchmarks evaluate action recognition on consumer videos, egocentric recordings, or simulated industrial environments. They do not test

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

Streaming Model Cascades for Semantic SQL

DGX agent

arXiv:2604.00660v2 Announce Type: replace-cross Abstract: Modern data warehouses extend SQL with semantic operators that invoke large language models on each qualifying row, making per-row inference o

model-releasesarxiv-cs-ai
7 Jul 2026
← Previous
1…3233343536…1021
Next →