AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,975 results
Model Releases

vLLM Semantic Router: Signal Driven Decision Routing for Mixture-of-Modality Models

DGX agent

arXiv:2603.04444v3 Announce Type: replace-cross Abstract: As large language models (LLMs) diversify across modalities, capabilities, and cost profiles, the problem of intelligent request routing -- se

model-releasesarxiv-cs-ai
3 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Wavelet Fourier Diffuser: Frequency-Aware Diffusion Model for Reinforcement Learning

DGX agent

arXiv:2509.19305v2 Announce Type: replace-cross Abstract: Diffusion probability models have shown significant promise in offline reinforcement learning by directly modeling trajectory sequences. Howev

model-releasesarxiv-cs-ai
3 Jun 2026
Safety

A Practical Upper Bound on Selection Bias Effects in Medical Prediction Models

DGX agent

arXiv:2606.00563v1 Announce Type: cross Abstract: Selection bias is a common and often unavoidable aspect of real-world data that challenges the generalizability of machine learning models. When model

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

Auditing Asset-Specific Preferences in Financial Large Language Models: Evidence from Bitcoin Representations and Portfolio Allocation

DGX agent

arXiv:2606.02528v1 Announce Type: cross Abstract: Large language models now power robo-advisors and trading agents, yet whether they carry built-in biases toward specific assets is largely untested. W

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Benchmarks for Vision-Language Models in Urban Perception Should Be Reliability-Aware and Negotiated

DGX agent

arXiv:2606.00871v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly used to generate structured descriptions of street-level imagery for tasks such as streetscape auditing

model-releasesarxiv-cs-ai
2 Jun 2026
Tutorials

Can Vision Language Models Learn Intuitive Physics from Interaction?

DGX agent

arXiv:2602.06033v2 Announce Type: replace Abstract: Pre-trained vision language models do not have good intuitions about the physical world. Recent work has shown that supervised fine-tuning can impro

tutorialsarxiv-cs-lg
2 Jun 2026
Research

GottBERT: a pure German Language Model

DGX agent

arXiv:2012.02110v2 Announce Type: replace Abstract: Pre-trained language models have significantly advanced natural language processing (NLP), especially with the introduction of BERT and its optimize

researcharxiv-cs-cl
2 Jun 2026
Model Releases

OmniEEG-Bench: A Standardized Evaluation Benchmark for EEG Foundation Models

DGX agent

arXiv:2606.00815v1 Announce Type: new Abstract: Electroencephalography (EEG) supports a variety of brain-computer interface (BCI) tasks ranging from brain-state monitoring to human-LLM interactions. E

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Parameter-Efficient Fine-Tuning of Large Pretrained Models for Instance Segmentation Tasks

DGX agent

arXiv:2606.01947v1 Announce Type: cross Abstract: Research and applications in artificial intelligence have recently shifted with the rise of large pretrained models, which deliver state-of-the-art re

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Persona Attack: Incremental Memory Injection Jailbreak Attack against Large Language Models

DGX agent

arXiv:2606.00150v1 Announce Type: cross Abstract: As Large Language Models evolve for user convenience, vulnerability to jailbreak attacks continues to be reported despite ongoing efforts in safety tr

model-releasesarxiv-cs-ai
2 Jun 2026
Hardware

PortBERT: Navigating the Depths of Portuguese Language Models

DGX agent

arXiv:2606.02100v1 Announce Type: new Abstract: Transformer models dominate modern NLP, but efficient, language-specific models remain scarce. In Portuguese, most focus on scale or accuracy, often neg

hardwarearxiv-cs-cl
2 Jun 2026
Model Releases

Prototype Transformer: Towards Language Model Architectures Interpretable by Design

DGX agent

arXiv:2602.11852v2 Announce Type: replace Abstract: While state-of-the-art language models (LMs) surpass most humans in certain domains, their reasoning remains largely opaque, reducing trust and incr

model-releasesarxiv-cs-ai
2 Jun 2026
Applications

Quantized Reasoning Models Think They Need to Think Longer, but They Do Not

DGX agent

arXiv:2606.00206v1 Announce Type: new Abstract: Post-training quantization (PTQ) is widely used to deploy large language models efficiently, but its effect on reasoning models is not well understood.

applicationsarxiv-cs-lg
2 Jun 2026
Local Ai

Query Circuits: Explaining How Language Models Answer User Prompts

DGX agent

arXiv:2509.24808v2 Announce Type: replace Abstract: Explaining why a language model produces a particular output requires local, input-level explanations. Existing methods uncover global capability ci

local-aiarxiv-cs-ai
2 Jun 2026
Model Releases

RoboTrustBench: Benchmarking the Trustworthiness of Video World Models for Robotic Manipulation

DGX agent

arXiv:2606.01600v1 Announce Type: cross Abstract: Video world models are increasingly used in robotic manipulation, yet existing benchmarks mostly evaluate them under valid, feasible, and safe instruc

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

The Refusal--Compliance Tradeoff: A Large-Scale Safety Behavior Audit of Large Language Models

DGX agent

arXiv:2605.05427v2 Announce Type: replace Abstract: Refusal rates are a poor proxy for LLM safety, i.e., a model may over-refuse benign prompts while still complying with harmful ones. We audit both f

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

WorldCache: Accelerating World Models for Free via Heterogeneous Token Caching

DGX agent

arXiv:2603.06331v2 Announce Type: replace Abstract: Diffusion-based world models have shown strong potential for unified world simulation, but the iterative denoising remains too costly for interactiv

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

WorldLens: Full-Spectrum Evaluations of Driving World Models in Real World

DGX agent

arXiv:2512.10958v2 Announce Type: replace Abstract: Generative world models are reshaping embodied AI, enabling agents to synthesize realistic 4D driving environments that look convincing but often fa

model-releasesarxiv-cs-cv
2 Jun 2026
Tutorials

Esoteric Language Models: A Family of Any-Order Diffusion LLMs

DGX agent

arXiv:2506.01928v4 Announce Type: replace Abstract: Diffusion-based language models offer a compelling alternative to autoregressive (AR) models by enabling parallel and controllable generation. Withi

tutorialsarxiv-cs-cl
1 Jun 2026
Model Releases

Exploring Autonomous Agentic Data Engineering for Model Specialization

DGX agent

arXiv:2605.30407v1 Announce Type: cross Abstract: Large Language Models (LLMs) have demonstrated strong performance on general tasks, while often struggling to adapt to specialized domains without hig

model-releasesarxiv-cs-ai
1 Jun 2026
Safety

Human-Alignment and Calibration of Inference-Time Uncertainty in Large Language Models

DGX agent

arXiv:2508.08204v2 Announce Type: replace-cross Abstract: There has been much recent interest in evaluating large language models for uncertainty calibration to facilitate model control and modulate u

safetyarxiv-cs-ai
1 Jun 2026
Model Releases

Quantifying the Uncertainty of Foundation Models with Singular Value Ensembles

DGX agent

arXiv:2601.22068v2 Announce Type: replace Abstract: Foundation models have become a dominant paradigm in machine learning, achieving remarkable performance across diverse tasks through large-scale pre

model-releasesarxiv-cs-lg
1 Jun 2026
Research

SALAAD: Sparse And Low-Rank Adaptation via ADMM for Large Language Model Inference

DGX agent

arXiv:2602.00942v3 Announce Type: replace Abstract: Modern large language models are increasingly deployed under compute and memory constraints, making flexible control of model capacity a central cha

researcharxiv-cs-lg
1 Jun 2026
Research

VLM3: Vision Language Models Are Native 3D Learners

DGX agent

arXiv:2605.30561v1 Announce Type: cross Abstract: Vision Language Models (VLMs) enable a unified model to solve various vision tasks through prompting. They have shown promising performance in semanti

researcharxiv-cs-ai
1 Jun 2026
Model Releases

When LLMs Learn to Be Consistently Wrong: A Multi-Model Study of Linear Representations of Synthetic Deception

DGX agent

arXiv:2605.30381v1 Announce Type: cross Abstract: Deceptive alignment, in which models maintain accurate internal representations while deliberately producing false outputs, remains a central challeng

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

A Minimal Bifurcation Model of Load Imbalance in a Softmax Mixture-of-Experts Router

DGX agent

arXiv:2605.29121v1 Announce Type: cross Abstract: We propose a minimal dynamical model of adaptive softmax routing for a two-expert Mixture-of-Experts (MoE) layer. The model is obtained as a mean-fiel

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Are LLMs Socially Adaptive? Contrasting Belief Evolution in Large Language Models and Humans

DGX agent

arXiv:2410.10398v3 Announce Type: replace-cross Abstract: As large language models (LLMs) increasingly engage in complex social interactions, ensuring that their behaviors align with human ethical pri

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Benchmarking Large Vision-Language Models on CFMME: A Comprehensive Chinese Financial Multimodal Evaluation Dataset

DGX agent

arXiv:2605.29462v1 Announce Type: cross Abstract: The emergence of Large Vision-Language Models (LVLMs) has substantially expanded model capabilities beyond text-only understanding, enabling unified i

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

CrystalXRD-Bench: Benchmarking Vision-Language Models for XRD Peak Indexing Across Diverse Crystalline Materials

DGX agent

arXiv:2605.29446v1 Announce Type: new Abstract: Miller-index identification from powder XRD patterns requires capabilities untested by existing multimodal benchmarks: the model must read a narrow peak

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

FarSkip-Collective: Unhobbling Blocking Communication in Mixture of Experts Models

DGX agent

arXiv:2511.11505v3 Announce Type: replace Abstract: Blocking communication presents a major hurdle in running MoEs efficiently in distributed settings. To address this, we present FarSkip-Collective w

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

KBF: Knowledge Boundary as Fingerprint for Language Model and Black-Box API Auditing

DGX agent

arXiv:2605.29524v1 Announce Type: cross Abstract: Relay and reseller APIs increasingly intermediate access to large language models (LLMs), but users have no direct way to verify that a claimed endpoi

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Kronecker Embeddings: Byte-Level Structured Token Representations for Parameter-Efficient Language Models

DGX agent

arXiv:2605.29459v1 Announce Type: new Abstract: Large language models route every input through a learned embedding table of shape |V| x d_model, consuming hundreds of millions to billions of trainabl

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

Leveraging Routing Dynamics in Mixture-of-Experts Models for Efficient Language Adaptation

DGX agent

arXiv:2605.29714v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models are widely used to scale language models, yet their expert routing behavior and adaptation in a multilingual setting rem

model-releasesarxiv-cs-cl
29 May 2026
Agents

MemCollab: Cross-Model Memory Collaboration via Contrastive Trajectory Distillation

DGX agent

arXiv:2603.23234v2 Announce Type: replace Abstract: LLM agents increasingly rely on memory mechanisms to reuse knowledge from past problem-solving experiences. However, existing methods typically cons

agentsarxiv-cs-ai
29 May 2026
Research

minWM: A Full-Stack Open-Source Framework for Real-Time Interactive Video World Models

DGX agent

arXiv:2605.30263v1 Announce Type: new Abstract: Recent video diffusion foundation models have achieved remarkable progress in high-quality video generation, yet turning them into real-time interactive

researcharxiv-cs-cv
29 May 2026
Safety

Rubric-Guided Process Reward for Stepwise Model Routing

DGX agent

arXiv:2605.29310v1 Announce Type: new Abstract: Stepwise model routing improves the efficiency of Large Reasoning Models (LRMs) by assigning each reasoning step to a suitable model. Recent methods for

safetyarxiv-cs-ai
29 May 2026
Model Releases

The Price Reversal Phenomenon: When Cheaper Reasoning Models Cost More

DGX agent

arXiv:2603.23971v2 Announce Type: replace-cross Abstract: Developers and consumers increasingly choose reasoning models (RMs) based on their listed API prices. However, how accurately do these prices

model-releasesarxiv-cs-ai
29 May 2026
Local Ai

Unsupervised Semantic Segmentation Facilitates Model Understanding

DGX agent

arXiv:2605.29691v1 Announce Type: new Abstract: Self-supervised learning (SSL) has produced a diverse landscape of vision transformers (ViTs) whose pretrained representations support a wide range of d

local-aiarxiv-cs-cv
29 May 2026
Agents

Adapting the Interface, Not the Model: Runtime Harness Adaptation for Deterministic LLM Agents

DGX agent

arXiv:2605.22166v2 Announce Type: replace Abstract: LLM agents are shaped not only by their language models, but also by the runtime harness that mediates observation, tool use, action execution, feed

agentsarxiv-cs-ai
28 May 2026
Safety

Bridging the Detection-to-Abstention Gap in Reasoning Models under Insufficient Information

DGX agent

arXiv:2605.28070v1 Announce Type: new Abstract: We highlight a failure mode of large reasoning models on questions with insufficient information: models may recognize that a problem is under-specified

safetyarxiv-cs-ai
28 May 2026
Model Releases

Can Segmentation Models Understand the World? Towards Proactive Affordance Reasoning via Visual Chain-of-Thought

DGX agent

arXiv:2605.27764v1 Announce Type: cross Abstract: Recent segmentation models couple large language models (LLMs) with mask decoders to ground complex language expressions into masks, yet their instruc

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

CogVLA: Cognition-Aligned Vision-Language-Action Model via Instruction-Driven Routing & Sparsification

DGX agent

arXiv:2508.21046v3 Announce Type: replace Abstract: Recent Vision-Language-Action (VLA) models built on pre-trained Vision-Language Models (VLMs) require extensive post-training, resulting in high com

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

Do Models Know Why They Changed Their Mind? Interpretability and Faithfulness of Chain-of-Thought Under Knowledge Conflict

DGX agent

arXiv:2605.27773v1 Announce Type: cross Abstract: When a language model sees a document contradicting its training knowledge, it must choose: follow the document or trust itself. Prior work proved thi

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

JMedEthicBench: A Multi-Turn Conversational Benchmark for Evaluating Medical Safety in Japanese Large Language Models

DGX agent

arXiv:2601.01627v3 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) are increasingly deployed in healthcare field, it becomes essential to carefully evaluate their medical safety

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

XTransfer: Modality-Agnostic Few-Shot Model Transfer for Human Sensing at the Edge

DGX agent

arXiv:2506.22726v4 Announce Type: replace Abstract: Deep learning for human sensing on edge systems presents significant potential for smart applications. However, its training and development are hin

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

AdaSD: Adaptive Speculative Decoding for Efficient Language Model Inference

DGX agent

arXiv:2512.11280v2 Announce Type: replace Abstract: Large language models (LLMs) have achieved remarkable performance across a wide range of tasks, but their increasing parameter sizes significantly s

model-releasesarxiv-cs-cl
27 May 2026
Safety

Alignment Makes Language Models Normative, Not Descriptive

DGX agent

arXiv:2603.17218v2 Announce Type: replace-cross Abstract: Post-training alignment optimizes language models to match human preference signals, but this objective is not equivalent to modeling observed

safetyarxiv-cs-ai
27 May 2026
Model Releases

Beyond Questions: Evaluating What Large Language Models (Actually) Know

DGX agent

arXiv:2605.26937v1 Announce Type: cross Abstract: Parametric knowledge in large language models (LLMs) is a cornerstone of their success, yet remains poorly understood. Existing knowledge benchmarks t

model-releasesarxiv-cs-ai
27 May 2026
← Previous
1…3536373839…1021
Next →