AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,975 results
Model Releases

GeMoE: Gating Entropy is All You Need for Uncertainty-aware Adaptive Routing in MoE-based Large Vision-Language Models

DGX agent

arXiv:2606.26287v1 Announce Type: new Abstract: With the increase in model parameters and training data, the instruction following and generalization capabilities of Large VisionLanguage Models (LVLMs

model-releasesarxiv-cs-cv
26 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Hallucination in World Models is Predictable and Preventable

DGX agent

arXiv:2606.27326v1 Announce Type: cross Abstract: Modern generative world models render increasingly realistic action-controllable futures, yet they frequently hallucinate: rollouts remain visually fl

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

HarmVideoBench: Benchmarking Harmful Video Understanding in Large Multimodal Models

DGX agent

arXiv:2606.27187v1 Announce Type: cross Abstract: Large vision-language models (LVLMs) have recently shown immense potential in automated content moderation, sparking growing interest in developing ha

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

Learning to Recover Task Experts from a Multi-Task Merged Model

DGX agent

arXiv:2606.26902v1 Announce Type: new Abstract: Multi-task model merging aims to consolidate several task-specific experts into a unified model, yet static merging consistently suffers from parameter

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

RSPC: A Benchmark for Modeling Stress and Psychiatric Conditions in Digitally Mediated Relationships using Psychiatrist Annotations

DGX agent

arXiv:2606.27247v1 Announce Type: new Abstract: In NLP, mental health conditions are often modeled as isolated phenomena, without interpersonal context. We use Reddit posts about long-distance relatio

model-releasesarxiv-cs-lg
26 Jun 2026
Research

Sampling sea state using a diffusion model

DGX agent

arXiv:2606.26389v1 Announce Type: cross Abstract: Sea state prediction is essential for operational maritime applications and coupled earth system modeling, yet current spectral wave models remain com

researcharxiv-cs-ai
26 Jun 2026
Model Releases

The Inattentional Gap: Task-Conditioned Language and Vision Models Omit the Safety-Critical Signals They Can Otherwise Report

DGX agent

arXiv:2606.26529v1 Announce Type: cross Abstract: AI safety is evaluated by how reliably a model detects the hazards it is told to find, yet accidents often arise from the hazard no one specified. We

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation

DGX agent

arXiv:2606.25476v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated remarkable performance across natural language processing tasks, yet their deployment in high-stakes appl

model-releasesarxiv-cs-cl
25 Jun 2026
Safety

DRM: Diffusion-based Reward Model With Step-wise Guidance

DGX agent

arXiv:2605.25661v2 Announce Type: replace Abstract: Current mainstream methods of aligning diffusion models with human preferences typically employ VLM-based reward models. However, these reward model

safetyarxiv-cs-cv
25 Jun 2026
Model Releases

Internal Data Repetition Destroys Language Models

DGX agent

arXiv:2606.24998v1 Announce Type: new Abstract: Language models are running out of high-quality training data, and even aggressively deduplicated corpora retain some amount of repetition. Earlier cont

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Model Forensics: Investigating Whether Concerning Behavior Reflects Misalignment

DGX agent

arXiv:2606.26071v1 Announce Type: new Abstract: A central goal of safety research is determining whether a model is misaligned. Prior work has largely focused on detecting concerning behavior. But beh

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Small edits, large models: How Wikipedia advocacy shapes LLM values

DGX agent

arXiv:2606.24890v1 Announce Type: new Abstract: Can a small group of volunteers shape how AI systems discuss animal welfare, just by editing Wikipedia? We show that they can. Wikipedia appears in near

model-releasesarxiv-cs-cl
25 Jun 2026
Local Ai

CrossPool: Efficient Multi-LLM Serving for Cold MoE Models through KV-Cache and Weight Disaggregation

DGX agent

arXiv:2606.24506v1 Announce Type: cross Abstract: Emerging LLM services increasingly host many sparse MoE models, yet most models receive sparse requests and remain cold. This creates a GPU memory pro

local-aiarxiv-cs-ai
24 Jun 2026
Model Releases

Experiments with Optimal Model Trees

DGX agent

arXiv:2503.12902v4 Announce Type: replace Abstract: Model trees provide an appealing way to perform interpretable machine learning for both classification and regression problems. In contrast to ``cla

model-releasesarxiv-cs-lg
24 Jun 2026
Model Releases

Grounded Chess Reasoning in Language Models via Master Distillation

DGX agent

arXiv:2603.20510v2 Announce Type: replace Abstract: Language models often lack grounded reasoning capabilities in specialized domains where training data is scarce but bespoke systems excel. We introd

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

OpenThoughts-Agent: Data Recipes for Agentic Models

DGX agent

arXiv:2606.24855v1 Announce Type: new Abstract: Agentic language models dramatically expand the applications of AI yet little is publicly known about how to curate training data for broadly capable ag

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

Rapid FinFET Modelling Using an Autoencoder

DGX agent

arXiv:2606.24046v1 Announce Type: cross Abstract: This work presents a machine learning framework that leverages an autoencoder (AE) for the efficient modeling of FinFET. We first calibrated a BSIM-CM

model-releasesarxiv-cs-ai
24 Jun 2026
Research

Beyond the Next Step: Variable-Length Latent World Models for Long-Horizon Planning

DGX agent

arXiv:2606.21775v1 Announce Type: new Abstract: Recently, world models have emerged as a promising paradigm for building intelligent agents by learning predictive models that estimate future environme

researcharxiv-cs-lg
23 Jun 2026
Research

Discretizing Reward Models

DGX agent

arXiv:2606.21795v1 Announce Type: new Abstract: Despite their widespread use, the role of reward models in shaping reinforcement learning is poorly understood. Reward models offer a tempting promise:

researcharxiv-cs-lg
23 Jun 2026
Model Releases

Oracle-RLAIF: An Improved Fine-Tuning Framework for Multi-modal Video Models using Reinforcement Learning from Ranking Feedback

DGX agent

arXiv:2510.02561v2 Announce Type: replace Abstract: Recent advances in large video-language models (VLMs) rely on extensive fine-tuning techniques that strengthen alignment between textual and visual

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Who Owns the AI Recommendation? A Multi-Industry Empirical Map of Brand Category Ownership Across Large Language Models

DGX agent

arXiv:2606.23057v1 Announce Type: cross Abstract: Large language models now mediate how buyers discover products and services, making the competitive structure of AI-generated recommendations a strate

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

ICA Lens: Interpreting Language Models Without Training Another Dictionary

DGX agent

arXiv:2606.11722v1 Announce Type: cross Abstract: Finding interpretable directions in language-model representations is critical for understanding and controlling model behavior. Sparse autoencoders (

model-releasesarxiv-cs-ai
11 Jun 2026
Applications

BiWM: Advancing Open-Source Interactive Video World Models with Bidirectional Autoregression

DGX agent

arXiv:2606.10135v1 Announce Type: cross Abstract: Transitioning bidirectional video diffusion models into an autoregressive paradigm improves the interactivity of video world models, but existing caus

applicationsarxiv-cs-ai
10 Jun 2026
Model Releases

Next Forcing: Causal World Modeling with Multi-Chunk Prediction

DGX agent

arXiv:2606.11187v1 Announce Type: new Abstract: Autoregressive video generation has emerged as a powerful paradigm for World Action Models (WAMs). However, existing approaches suffer from slow trainin

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

Calibration of Structured Ignorance Certificates for Diagnosing Unknown Unknowns in Reasoning Models

DGX agent

arXiv:2606.08571v1 Announce Type: cross Abstract: Large language models frequently fail in a characteristic way: rather than acknowledging ignorance, they produce fluent but incorrect answers to quest

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Component Ablation for Efficient Hybrid Language Model Architectures: Performance, Resilience, and Compression Implications

DGX agent

arXiv:2603.22473v2 Announce Type: replace-cross Abstract: Hybrid language models combine softmax attention with linear-time sequence mechanisms such as state-space or linear-attention layers, but the

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

DisCo: World Models with Discrete Camera Motion Control

DGX agent

arXiv:2606.07967v1 Announce Type: new Abstract: Controllable video world models target interactive world exploration, where models must faithfully execute explicit action commands while preserving vis

model-releasesarxiv-cs-cv
9 Jun 2026
Tutorials

From inverse problems to neural operators: prediction, mechanism, and generalization of data-driven models

DGX agent

arXiv:2606.08956v1 Announce Type: new Abstract: Scientists have historically relied on mathematical models based on differential equations to relate system inputs -- forces, fluxes, or heat sources --

tutorialsarxiv-cs-lg
9 Jun 2026
Model Releases

How Small Can You Go? LoRA Fine-Tuning 270M-8B Models for Merchant Information Extraction in Financial Transactions

DGX agent

arXiv:2606.08051v1 Announce Type: new Abstract: Financial transaction processing requires extracting structured merchant information from noisy, abbreviated bank transaction strings at scale. Our curr

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

IDDM: Identity-Decoupled Personalized Diffusion Models with a Tunable Privacy-Utility Trade-off

DGX agent

arXiv:2604.00903v2 Announce Type: replace Abstract: Personalized text-to-image diffusion models (e.g., DreamBooth, LoRA) enable users to synthesize high-fidelity avatars from a few reference photos fo

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Lost in the Non-convex Loss Landscape: How to Fine-tune the Large Time Series Model?

DGX agent

arXiv:2606.08578v1 Announce Type: new Abstract: Recently, large time series models (LTSMs) have gained increasing attention due to their similarities to large language models, including flexible conte

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

PRISM: PRior-guided Imagination Sampling in world Models

DGX agent

arXiv:2606.07974v1 Announce Type: cross Abstract: A learned world model provides a powerful physical intuition for evaluating future states. But its effectiveness in continuous control also depends cr

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

When Do Local Score Models Extrapolate Across Size? A Diagnostic Theory and Benchmark

DGX agent

arXiv:2606.09705v1 Announce Type: new Abstract: Scientific generative modeling often requires size transfer, where models trained on small systems are evaluated on larger ones. While translation-invar

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Where Instruction Hierarchy Breaks: Diagnosing and Repairing Failures in Reasoning Language Models

DGX agent

arXiv:2606.07808v1 Announce Type: new Abstract: Reasoning language models deployed in agentic workflows must follow an instruction hierarchy: when instructions from different sources conflict, the mod

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Closed-Form Spectral Regularization for Multi-Task Model Merging

DGX agent

arXiv:2606.07289v1 Announce Type: cross Abstract: Model merging combines several independently fine-tuned experts into a single multi-task model without any training data, reducing the storage, servin

model-releasesarxiv-cs-cv
8 Jun 2026
Research

Drifting Models for Surrogate Flow Modeling

DGX agent

arXiv:2606.07481v1 Announce Type: new Abstract: While Computational Fluid Dynamics (CFD) provides high-fidelity flow fields for optimizing indoor environments, its computational cost limits rapid expl

researcharxiv-cs-lg
8 Jun 2026
Research

Position: A Dynamical Systems Perspective is Needed to Advance Time Series Modeling

DGX agent

arXiv:2602.16864v2 Announce Type: replace-cross Abstract: Time series (TS) modeling has come a long way from early statistical, mainly linear, approaches to the current trend in TS foundation models.

researcharxiv-cs-ai
8 Jun 2026
Model Releases

Evaluating Large Language Models in Dynamic Clinical Decision-Making with Standardized Patient Cases

DGX agent

arXiv:2606.05112v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly proposed as clinical agents, yet static, single-turn benchmarks cannot capture how a model dynamically del

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

GENEB: Why Genomic Models Are Hard to Compare

DGX agent

arXiv:2606.04525v1 Announce Type: new Abstract: Progress in genomic foundation models is difficult to assess due to fragmented benchmarks, incompatible evaluation protocols, and task-specific reportin

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

Learning When to Act or Refuse: Guarding Agentic Reasoning Models for Safe Multi-Step Tool Use

DGX agent

arXiv:2603.03205v2 Announce Type: replace Abstract: Agentic language models operate in a fundamentally different safety regime than chat models: they must plan, call tools, and execute long-horizon ac

model-releasesarxiv-cs-cl
4 Jun 2026
Research

Conditional Latent Diffusion Model with Fourier-based Motion Modelling for Virtual Population Synthesis

DGX agent

arXiv:2606.03827v1 Announce Type: cross Abstract: In-silico trials of medical devices require the generation of virtual populations of anatomies. In cardiovascular applications, virtual anatomy is typ

researcharxiv-cs-ai
3 Jun 2026
Model Releases

Echelon: Auditable Aggregate-Only Language-Model Adaptation Across Privacy Boundaries

DGX agent

arXiv:2606.02958v1 Announce Type: cross Abstract: Cross-organization language-model adaptation increasingly faces hard governance constraints: in many deployments, device-level model state-parameters,

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

From Answers to States: Verifiable Process-Level Evaluation of Chemical Reasoning in Large Language Models

DGX agent

arXiv:2606.03660v1 Announce Type: new Abstract: Large language models are increasingly used as chemistry assistants, yet most chemistry benchmarks still score only final answers. This masks a critical

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Greener Than Humans? Environmental Attitudes in Large Language Models

DGX agent

arXiv:2606.02741v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used in sustainability-related decision support, reporting, and public communication, yet little systemati

model-releasesarxiv-cs-cl
3 Jun 2026
Agents

Comprehensive AI governance requires addressing non-model gains

DGX agent

arXiv:2606.00047v1 Announce Type: cross Abstract: Frontier AI governance often centres on the model-level governance paradigm, which assumes that a model's capability profile is primarily a function o

agentsarxiv-cs-ai
2 Jun 2026
Safety

Emergent Collaborative Deliberation in Multi-Model AI Systems: A BFT-Derived Protocol for Epistemic Synthesis

DGX agent

arXiv:2606.00005v1 Announce Type: new Abstract: We present the Consilium Protocol, a Byzantine Fault Tolerance-derived architecture for structured multi-model AI deliberation that treats inter-model d

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

Geometry-Aware Implicit Memory for Video World Models

DGX agent

arXiv:2606.02436v1 Announce Type: new Abstract: Video world models aim to simulate controllable visual environments, but long-horizon rollouts depend on what the model remembers after observations lea

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

MBench: A Comprehensive Benchmark on Memory Capability for Video World Models

DGX agent

arXiv:2606.00793v1 Announce Type: new Abstract: Recent advancements in video-based world models have demonstrated an unprecedented ability to synthesize high-fidelity visual sequences. However, a fund

model-releasesarxiv-cs-cv
2 Jun 2026
← Previous
1…2425262728…1021
Next →