AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,587 results
Research

Teaching Language Models Mechanistic Explainability Through MechSMILES

DGX agent

arXiv:2512.05722v2 Announce Type: replace Abstract: Chemical reaction mechanisms are the foundation of how chemists evaluate reactivity and feasibility, yet current Computer-Assisted Synthesis Plannin

researcharxiv-cs-lg
20 Apr 2026
Hardware
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

The threat of analytic flexibility in using large language models to simulate human data

DGX agent

arXiv:2509.13397v3 Announce Type: replace-cross Abstract: Social scientists are now using large language models to create 'silicon samples': synthetic datasets intended to stand in for human responden

hardwarearxiv-cs-ai
20 Apr 2026
Research

VIB-Probe: Detecting and Mitigating Hallucinations in Vision-Language Models via Variational Information Bottleneck

DGX agent

arXiv:2601.05547v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) have demonstrated remarkable progress in multimodal tasks, but remain susceptible to hallucinations, where gener

researcharxiv-cs-ai
20 Apr 2026
Applications

A major lesson to take away from Opus 4.7 is that, while there is a lot of arguments about implementation choices and personality, models ke…

DGX agent

A major lesson to take away from Opus 4.7 is that, while there is a lot of arguments about implementation choices and personality, models keep improving measurably on economically important tasks with

applicationsethan-mollick--x
18 Apr 2026
Research

AlphaCNOT: Learning CNOT Minimization with Model-Based Planning

DGX agent

arXiv:2604.13812v1 Announce Type: new Abstract: Quantum circuit optimization is a central task in Quantum Computing, as current Noisy Intermediate Scale Quantum devices suffer from error propagation t

researcharxiv-cs-ai
17 Apr 2026
Model Releases

AnimationBench: Are Video Models Good at Character-Centric Animation?

DGX agent

arXiv:2604.15299v1 Announce Type: new Abstract: Video generation has advanced rapidly, with recent methods producing increasingly convincing animated results. However, existing benchmarks-largely desi

model-releasesarxiv-cs-cv
17 Apr 2026
Research

Bayesian-LoRA: Probabilistic Low-Rank Adaptation of Large Language Models

DGX agent

arXiv:2601.21003v2 Announce Type: replace Abstract: Large Language Models usually put more emphasis on accuracy and therefore, will guess even when not certain about the prediction, which is especiall

researcharxiv-cs-ai
17 Apr 2026
Research

Compressing Sequences in the Latent Embedding Space: K-Token Merging for Large Language Models

DGX agent

arXiv:2604.15153v1 Announce Type: new Abstract: Large Language Models (LLMs) incur significant computational and memory costs when processing long prompts, as full self-attention scales quadratically

researcharxiv-cs-cl
17 Apr 2026
Research

Dark & Stormy: Modeling Humor in Sentences from the Bulwer-Lytton Fiction Contest

DGX agent

arXiv:2510.24538v2 Announce Type: replace Abstract: Textual humor is enormously diverse and computational studies need to account for this range, including intentionally bad humor. In this paper, we c

researcharxiv-cs-cl
17 Apr 2026
Safety

Exploration and Exploitation Errors Are Measurable for Language Model Agents

DGX agent

arXiv:2604.13151v1 Announce Type: new Abstract: Language Model (LM) agents are increasingly used in complex open-ended decision-making tasks, from AI coding to physical AI. A core requirement in these

safetyarxiv-cs-ai
17 Apr 2026
Model Releases

ImplicitMemBench: Measuring Unconscious Behavioral Adaptation in Large Language Models

DGX agent

arXiv:2604.08064v2 Announce Type: replace Abstract: Existing memory benchmarks for LLM agents evaluate explicit recall of facts, yet overlook implicit memory where experience becomes automated behavio

model-releasesarxiv-cs-ai
17 Apr 2026
Applications

IUQ: Interrogative Uncertainty Quantification for Long-Form Large Language Model Generation

DGX agent

arXiv:2604.15109v1 Announce Type: new Abstract: Despite the rapid advancement of Large Language Models (LLMs), uncertainty quantification in LLM generation is a persistent challenge. Although recent a

applicationsarxiv-cs-cl
17 Apr 2026
Local Ai

Keep It CALM: Toward Calibration-Free Kilometer-Level SLAM with Visual Geometry Foundation Models via an Assistant Eye

DGX agent

arXiv:2604.14795v1 Announce Type: new Abstract: Visual Geometry Foundation Models (VGFMs) demonstrate remarkable zero-shot capabilities in local reconstruction. However, deploying them for kilometer-l

local-aiarxiv-cs-ro
17 Apr 2026
Safety

Model-Free Assessment of Simulator Fidelity via Quantile Curves

DGX agent

arXiv:2512.05024v3 Announce Type: replace-cross Abstract: As generative AI models are increasingly used to simulate real-world systems, quantifying the ``sim-to-real'' gap is critical. For each input

safetyarxiv-cs-lg
17 Apr 2026
Safety

Modeling LLM Unlearning as an Asymmetric Two-Task Learning Problem

DGX agent

arXiv:2604.14808v1 Announce Type: new Abstract: Machine unlearning for large language models (LLMs) aims to remove targeted knowledge while preserving general capability. In this paper, we recast LLM

safetyarxiv-cs-cl
17 Apr 2026
Safety

The PICCO Framework for Large Language Model Prompting: A Taxonomy and Reference Architecture for Prompt Structure

DGX agent

arXiv:2604.14197v1 Announce Type: new Abstract: Large language model (LLM) performance depends heavily on prompt design, yet prompt construction is often described and applied inconsistently. Our purp

safetyarxiv-cs-cl
17 Apr 2026
Applications

Alibaba's new Token Hub unit releases Happy Oyster, a new AI world model that can create 3D environments, interactive videos, films, video content, and games (Luz Ding/Bloomberg)

DGX agent

Luz Ding / Bloomberg: Alibaba's new Token Hub unit releases Happy Oyster, a new AI world model that can create 3D environments, interactive videos, films, video content, and games — Alibaba Group Hold

applicationstechmeme
16 Apr 2026
Safety

Character Beyond Speech: Leveraging Role-Playing Evaluation in Audio Large Language Models via Reinforcement Learning

DGX agent

arXiv:2604.13804v1 Announce Type: new Abstract: The rapid evolution of multimodal large models has revolutionized the simulation of diverse characters in speech dialogue systems, enabling a novel inte

safetyarxiv-cs-lg
16 Apr 2026
Agents

Co-FactChecker: A Framework for Human-AI Collaborative Claim Verification Using Large Reasoning Models

DGX agent

arXiv:2604.13706v1 Announce Type: new Abstract: Professional fact-checkers rely on domain knowledge and deep contextual understanding to verify claims. Large language models (LLMs) and large reasoning

agentsarxiv-cs-cl
16 Apr 2026
Research

Don't Let the Video Speak: Audio-Contrastive Preference Optimization for Audio-Visual Language Models

DGX agent

arXiv:2604.14129v1 Announce Type: new Abstract: While Audio-Visual Language Models (AVLMs) have achieved remarkable progress over recent years, their reliability is bottlenecked by cross-modal halluci

researcharxiv-cs-cv
16 Apr 2026
Research

Empirical Evidence of Complexity-Induced Limits in Large Language Models on Finite Discrete State-Space Problems with Explicit Validity Constraints

DGX agent

arXiv:2604.13371v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly described as possessing strong reasoning capabilities, supported by high performance on mathematical, logi

researcharxiv-cs-cl
16 Apr 2026
Safety

Foresight Optimization for Strategic Reasoning in Large Language Models

DGX agent

arXiv:2604.13592v1 Announce Type: new Abstract: Reasoning capabilities in large language models (LLMs) have generally advanced significantly. However, it is still challenging for existing reasoning-ba

safetyarxiv-cs-cl
16 Apr 2026
Local Ai

From Anchors to Supervision: Memory-Graph Guided Corpus-Free Unlearning for Large Language Models

DGX agent

arXiv:2604.13777v1 Announce Type: new Abstract: Large language models (LLMs) may memorize sensitive or copyrighted content, raising significant privacy and legal concerns. While machine unlearning has

local-aiarxiv-cs-cl
16 Apr 2026
Model Releases

GeoBridge: A Semantic-Anchored Multi-View Foundation Model Bridging Images and Text for Geo-Localization

DGX agent

arXiv:2512.02697v3 Announce Type: replace Abstract: Cross-view geo-localization infers a location by retrieving geo-tagged reference images that visually correspond to a query image. However, the trad

model-releasesarxiv-cs-cv
16 Apr 2026
Local Ai

Great blogpost from @pcuenq on making a new skill + test harness to automate porting new models from Transformers to mlx-lm

DGX agent

This post discusses a blog article by @pcuenq that covers the process of creating new skills and test harnesses to automate the conversion of machine learning models from the Hugging Face Transformers

local-aiclem-delangue--x
16 Apr 2026
Research

How Can We Synthesize High-Quality Pretraining Data? A Systematic Study of Prompt Design, Generator Model, and Source Data

DGX agent

arXiv:2604.13977v1 Announce Type: new Abstract: Synthetic data is a standard component in training large language models, yet systematic comparisons across design dimensions, including rephrasing stra

researcharxiv-cs-cl
16 Apr 2026
Industry

Introducing GPT-Rosalind, our frontier reasoning model built to support research across biology, drug discovery, and translational medicine.

DGX agent

GPT-Rosalind is a frontier reasoning model developed by OpenAI, specifically designed to support research applications in biology, drug discovery, and translational medicine. The model was introduced

industrysam-altman--x
16 Apr 2026
Research

Linear Probe Accuracy Scales with Model Size and Benefits from Multi-Layer Ensembling

DGX agent

arXiv:2604.13386v1 Announce Type: new Abstract: Linear probes can detect when language models produce outputs they 'know' are wrong, a capability relevant to both deception and reward hacking. However

researcharxiv-cs-lg
16 Apr 2026
Tutorials

Modeling Student Learning with 3.8 Million Program Traces

DGX agent

arXiv:2510.05056v2 Announce Type: replace Abstract: As programmers write code, they often edit and retry multiple times, creating rich 'interaction traces' that reveal how they approach coding tasks a

tutorialsarxiv-cs-lg
16 Apr 2026
Research

Native Hybrid Attention for Efficient Sequence Modeling

DGX agent

arXiv:2510.07019v3 Announce Type: replace Abstract: Transformers excel at sequence modeling but face quadratic complexity, while linear attention offers improved efficiency but often compromises recal

researcharxiv-cs-cl
16 Apr 2026
Research

Quantifying and Understanding Uncertainty in Large Reasoning Models

DGX agent

arXiv:2604.13395v1 Announce Type: cross Abstract: Large Reasoning Models (LRMs) have recently demonstrated significant improvements in complex reasoning. While quantifying generation uncertainty in LR

researcharxiv-cs-lg
16 Apr 2026
Research

Sign up today @ http://portal.nousresearch.com/manage-subscription Run `hermes update` and then `hermes model` to configure

DGX agent

Nous Research is promoting their subscription portal and providing instructions for users to update and configure their Hermes model through command-line tools (`hermes update` and `hermes model`). Th

researchnous-research--x
16 Apr 2026
Research

CLASP: Class-Adaptive Layer Fusion and Dual-Stage Pruning for Multimodal Large Language Models

DGX agent

arXiv:2604.12767v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) suffer from substantial computational overhead due to the high redundancy in visual token sequences. Existing

researcharxiv-cs-ai
15 Apr 2026
Research

Cognition-Inspired Dual-Stream Semantic Enhancement for Vision-Based Dynamic Emotion Modeling

DGX agent

arXiv:2604.12777v1 Announce Type: cross Abstract: The human brain constructs emotional percepts not by processing facial expressions in isolation, but through a dynamic, hierarchical integration of se

researcharxiv-cs-ai
15 Apr 2026
Research

CycloneMAE: A Scalable Multi-Task Learning Model for Global Tropical Cyclone Probabilistic Forecasting

DGX agent

arXiv:2604.12180v1 Announce Type: cross Abstract: Tropical cyclones (TCs) rank among the most destructive natural hazards, yet their forecasting faces fundamental trade-offs: numerical weather predict

researcharxiv-cs-ai
15 Apr 2026
Research

Efficient Inference for Large Vision-Language Models: Bottlenecks, Techniques, and Prospects

DGX agent

arXiv:2604.05546v2 Announce Type: replace Abstract: Large Vision-Language Models (LVLMs) enable sophisticated reasoning over images and videos, yet their inference is hindered by a systemic efficiency

researcharxiv-cs-cl
15 Apr 2026
Local Ai

Evaluating Language Models for Harmful Manipulation

DGX agent

arXiv:2603.25326v4 Announce Type: replace Abstract: Interest in the concept of AI-driven harmful manipulation is growing, yet current approaches to evaluating it are limited. This paper introduces a f

local-aiarxiv-cs-ai
15 Apr 2026
Model Releases

How memory can affect collective and cooperative behaviors in an LLM-Based Social Particle Swarm

DGX agent

arXiv:2604.12250v1 Announce Type: new Abstract: This study examines how model-specific characteristics of Large Language Model (LLM) agents, including internal alignment, shape the effect of memory on

model-releasesarxiv-cs-ai
15 Apr 2026
Local Ai

How much more useage do you get from a $20 pro plan when using cloud models? Or OpenRouter better??

DGX agent

This Reddit thread from r/ollama discusses the value comparison between Ollama's 20/month Pro plan for cloud model usage versus using OpenRouter as an alternative. Ollama Cloud offers fixed-price subs

local-air-ollama
15 Apr 2026
Safety

InsightFlow: LLM-Driven Synthesis of Patient Narratives for Mental Health into Causal Models

DGX agent

arXiv:2604.12721v1 Announce Type: new Abstract: Clinical case formulation organizes patient symptoms and psychosocial factors into causal models, often using the 5P framework. However, constructing su

safetyarxiv-cs-cl
15 Apr 2026
Model Releases

LLM-Enhanced Log Anomaly Detection: A Comprehensive Benchmark of Large Language Models for Automated System Diagnostics

DGX agent

arXiv:2604.12218v1 Announce Type: new Abstract: System log anomaly detection is critical for maintaining the reliability of large-scale software systems, yet traditional methods struggle with the hete

model-releasesarxiv-cs-lg
15 Apr 2026
Local Ai

Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large Language Models

DGX agent

arXiv:2601.14004v4 Announce Type: replace Abstract: Mechanistic Interpretability (MI) has emerged as a vital approach to demystify the opaque decision-making of Large Language Models (LLMs). However,

local-aiarxiv-cs-cl
15 Apr 2026
Applications

Mantis: A Foundation Model for Mechanistic Disease Forecasting

DGX agent

arXiv:2508.12260v5 Announce Type: replace Abstract: Infectious disease forecasting in novel outbreaks or low-resource settings is hampered by the need for large disease and covariate data sets, bespok

applicationsarxiv-cs-ai
15 Apr 2026
Industry

Microsoft’s MAI-Image-2-Efficient model accelerates company’s move away from OpenAI

DGX agent

Microsoft Corp.’s push for artificial intelligence independence is gaining traction with today’s release of MAI-Image-2-Efficient, a lean and mean version of its flagship image generation model that d

industrysiliconangle
15 Apr 2026
Model Releases

ParetoBandit: Budget-Paced Adaptive Routing for Non-Stationary LLM Serving

DGX agent

arXiv:2604.00136v2 Announce Type: replace-cross Abstract: Multi-model LLM serving operates in a non-stationary, noisy environment: providers revise pricing, model quality can shift or regress without

model-releasesarxiv-cs-cl
15 Apr 2026
Safety

Preventing Safety Drift in Large Language Models via Coupled Weight and Activation Constraints

DGX agent

arXiv:2604.12384v1 Announce Type: new Abstract: Safety alignment in Large Language Models (LLMs) remains highly fragile during fine-tuning, where even benign adaptation can degrade pre-trained refusal

safetyarxiv-cs-ai
15 Apr 2026
Model Releases

Revisiting the Reliability of Language Models in Instruction-Following

DGX agent

arXiv:2512.14754v2 Announce Type: replace-cross Abstract: Advanced LLMs have achieved near-ceiling instruction-following accuracy on benchmarks such as IFEval. However, these impressive scores do not

model-releasesarxiv-cs-ai
15 Apr 2026
Applications

Robotic Manipulation is Vision-to-Geometry Mapping (f(v) rightarrow G): Vision-Geometry Backbones over Language and Video Models

DGX agent

arXiv:2604.12908v1 Announce Type: new Abstract: At its core, robotic manipulation is a problem of vision-to-geometry mapping (f(v) rightarrow G). Physical actions are fundamentally defined by geometri

applicationsarxiv-cs-ro
15 Apr 2026
← Previous
1…179180181182183…1263
Next →