AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,965
  • Agents7,446
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,131
  • Local Ai4,857
  • Model Releases23,360
  • Research19,834
  • Safety13,174
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,965
  • Agents7,446
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,131
  • Local Ai4,857
  • Model Releases23,360
  • Research19,834
  • Safety13,174
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
86,965Total entries
1Added by human
86,964Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,106 results
Model Releases

Predictions as Surrogates: Revisiting Surrogate Outcomes in the Age of AI

DGX agent

arXiv:2501.09731v2 Announce Type: replace-cross Abstract: We establish a formal connection between the decades-old surrogate outcome model in biostatistics and economics and the emerging field of pred

model-releasesarxiv-cs-lg
23 Jun 2026
Research
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Probabilistic Retrofitting of Learned Simulators

DGX agent

arXiv:2603.01949v2 Announce Type: replace Abstract: Dominant approaches for modelling Partial Differential Equations (PDEs) rely on deterministic predictions, yet many physical systems of interest are

researcharxiv-cs-lg
23 Jun 2026
Model Releases

ReNIO: Reweighting Negative Trajectory Importance for LLM On-Policy Distillation

DGX agent

arXiv:2606.23104v1 Announce Type: new Abstract: On-policy distillation (OPD) improves LLM reasoning by training a student model on its own generated outputs, but standard OPD treats all student-genera

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

RLM-Cascade: Response-Level Speculative Decoding for Cost-Efficient LLM API Serving

DGX agent

arXiv:2606.22840v1 Announce Type: new Abstract: We present RLM-Cascade, a proxy-layer system that applies speculative decoding at the response level to reduce LLM API costs without requiring model arc

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

RouteJudge: An Open Platform for Reproducible and Preference-Aware LLM Routing

DGX agent

arXiv:2606.18774v2 Announce Type: replace Abstract: We present RouteJudge, an online pairwise preference evaluation framework for LLM routing systems, with a public platform available at https://route

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Sequential Minimal Optimization Algorithm for One-Class Support Vector Machines With Privileged Information

DGX agent

arXiv:2606.22210v1 Announce Type: new Abstract: One of the powerful techniques in data modeling is accounting for features that are available at the training stage, but are not available when the trai

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Set-based v.s. Distribution-based Representations of Epistemic Uncertainty: A Comparative Study

DGX agent

arXiv:2602.22747v2 Announce Type: replace Abstract: Epistemic uncertainty in neural networks is commonly modeled using two second-order paradigms: distribution-based representations, which rely on pos

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

TeleStyle V2: Beyond Content-Preserving Style Transfer with Self-Distillation and Distribution-Matching-Distillation

DGX agent

arXiv:2606.20709v1 Announce Type: new Abstract: Given a content reference and a style reference, content-preserving style transfer requires the model to generate stylized outputs with content and styl

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Towards Error-Free Long Video Generation

DGX agent

arXiv:2606.22370v1 Announce Type: new Abstract: Recent advances in video generation have made minute-level synthesis possible; however, generating long videos remains challenging due to error accumula

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Understanding Parallel Samplers in Masked Diffusion via Random Walks on Graphs

DGX agent

arXiv:2606.22976v1 Announce Type: new Abstract: In this paper, we propose using random walks on graphs as a verifiable sandbox to study different parallel sampling strategies in masked diffusion model

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

UniRank: Unified Rank Allocation for Low-Rank LLM Compression

DGX agent

arXiv:2606.21847v1 Announce Type: new Abstract: Low-rank decomposition serves as a promising compression paradigm for large language models, however, rank allocation remains challenging: manual rules

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

A Lightweight Multi-Agent Framework for Automated Concrete Barrier Design

DGX agent

arXiv:2606.12040v1 Announce Type: new Abstract: The design of reinforced concrete highway barriers is a safety-critical process that requires strict compliance with regulatory provisions such as the A

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Categorical Prior Lock-in: Why In-Context Learning Fails for Structured Data

DGX agent

arXiv:2606.11961v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as conditional generators for structured data, relying on in-context learning (ICL) to adapt to new

model-releasesarxiv-cs-ai
11 Jun 2026
Safety

From Consumption to Reflection: Designing Human-AI Relations for Stable Reasoning

DGX agent

arXiv:2606.11195v1 Announce Type: cross Abstract: Large language models (LLMs) have transformed how humans access information, but not how we reason with it. Their fluency accelerates consumption whil

safetyarxiv-cs-ai
11 Jun 2026
Hardware

INFRAMIND: Infrastructure-Aware Multi-Agent Orchestration

DGX agent

arXiv:2606.11440v1 Announce Type: new Abstract: Existing multi-agent LLM orchestration methods, ranging from brute-force ensembles to learned routers, select models and topologies based on task and mo

hardwarearxiv-cs-ai
11 Jun 2026
Model Releases

Neural-Parameterized Cellular Automata for Wildfire Spread

DGX agent

arXiv:2606.11676v1 Announce Type: cross Abstract: Traditional wildfire models rely on rigid, low-dimensional parameters and static fuel maps, frequently underpredicting fire spread. To address this we

model-releasesarxiv-cs-lg
11 Jun 2026
Research

On Subquadratic Architectures: From Applications to Principles

DGX agent

arXiv:2606.12364v1 Announce Type: new Abstract: Transformers dominate modern sequence modeling, but their quadratic attention incurs substantial computational cost. Subquadratic architectures offer a

researcharxiv-cs-lg
11 Jun 2026
Model Releases

RankVR: Low-Rank Structure Perception and Value Recalibration for Robust Composed Image Retrieval

DGX agent

arXiv:2606.11689v1 Announce Type: new Abstract: Composed Image Retrieval (CIR) constitutes a pivotal paradigm requiring models to perform joint reasoning on reference images and modification texts. Ho

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

SPEAR: A System for Post-Quantization Error-Adaptive Recovery Enabling Efficient Low-Bit LLM Serving

DGX agent

arXiv:2606.11244v1 Announce Type: cross Abstract: Efficient large language model (LLM) serving is increasingly constrained by deployment cost. Quantization is a key technique for reducing serving cost

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

STEAM: Squeeze and Transform Enhanced Attention Module

DGX agent

arXiv:2412.09023v3 Announce Type: replace Abstract: Channel and spatial attention mechanisms introduced in earlier work enhance the representational capabilities of deep convolutional neural networks

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

TAHOE: Text-to-SQL with Automated Hint Optimization from Experience

DGX agent

arXiv:2606.12387v1 Announce Type: cross Abstract: Large Language Models (LLMs) have democratized database access through Text-to-SQL, but moving from prototypes to production remains difficult. Real d

model-releasesarxiv-cs-ai
11 Jun 2026
Applications

Unifying Learning Dynamics and Generalization in Transformers Scaling Law

DGX agent

arXiv:2512.22088v3 Announce Type: replace-cross Abstract: The scaling law, a cornerstone of Large Language Model (LLM) development, predicts improvements in model performance with increasing computati

applicationsarxiv-cs-ai
11 Jun 2026
Model Releases

Visualizing LLM Latent Space Geometry Through Dimensionality Reduction

DGX agent

arXiv:2511.21594v3 Announce Type: replace Abstract: Large language models (LLMs) achieve state-of-the-art results across many natural language tasks, but their internal mechanisms remain difficult to

model-releasesarxiv-cs-lg
11 Jun 2026
Model Releases

When Generic Prompt Improvements Hurt: Evaluation-Driven Iteration for LLM Applications

DGX agent

arXiv:2601.22025v2 Announce Type: replace-cross Abstract: Evaluating Large Language Model (LLM) applications differs from conventional software testing because outputs are probabilistic, semantically

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Achieving Cloud-Grade SLOs for Local Mixture-of-Experts Inference through CPU-GPU Hybrid Design

DGX agent

arXiv:2606.10493v1 Announce Type: cross Abstract: Local deployment of large Mixture-of-Experts (MoE) models falls short of the service quality achieved in cloud-scale environments, even under low-conc

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

ASyMOB: Algebraic Symbolic Mathematical Operations Benchmark

DGX agent

arXiv:2505.23851v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly applied to symbolic mathematics, yet existing evaluations often conflate pattern memorization wi

model-releasesarxiv-cs-ai
10 Jun 2026
Research

AuRA: Internalizing Audio Understanding into LLMs as LoRA

DGX agent

arXiv:2606.11033v1 Announce Type: cross Abstract: Recent efforts to extend large language models (LLMs) to speech inputs typically rely on cascaded ASR-LLM pipelines, end-to-end speech-language models

researcharxiv-cs-ai
10 Jun 2026
Model Releases

Beyond APIs: Probing the Limits of MLLMs in Physical Tool Use

DGX agent

arXiv:2606.10803v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) excel at utilizing digital APIs and increasingly serve as the 'brain' of embodied AI, instructing robots to i

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

CoCoSI: Collaborative Cognitive Map Construction for Spatial Intelligence

DGX agent

arXiv:2606.10401v1 Announce Type: new Abstract: Spatial intelligence is a key frontier for multimodal large language models (MLLMs), enabling them to reason about the physical world from visual experi

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

ComBench: A Benchmark for Rigorous Proof Reasoning and Constructive Realization in Olympiad-Level Combinatorics

DGX agent

arXiv:2606.10479v1 Announce Type: new Abstract: Combinatorics is central to Olympiad-level mathematical problem solving, requiring deep discrete reasoning, creative constructions, and rigorous structu

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

CommonLID: Re-evaluating State-of-the-Art Language Identification Performance on Web Data

DGX agent

arXiv:2601.18026v2 Announce Type: replace Abstract: Language identification (LID) is a fundamental step in curating multilingual corpora. However, LID models still perform poorly for many languages, e

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

Compile Once, Differentiate Everywhere: A Differentiable Meta-Circular Interpreter

DGX agent

arXiv:2606.09930v1 Announce Type: cross Abstract: The boundary between program execution and gradient-based optimization has long limited the use of code itself as a learnable scientific model. We pre

model-releasesarxiv-cs-lg
10 Jun 2026
Research

Cost-Aware Routing for Efficient Text-To-Image Generation

DGX agent

arXiv:2506.14753v3 Announce Type: replace Abstract: Diffusion models are well known for their ability to generate a high-fidelity image for an input prompt through an iterative denoising process. Unfo

researcharxiv-cs-cv
10 Jun 2026
Model Releases

Expert-Level Crisis Detection in Mental Health Conversations

DGX agent

arXiv:2606.10380v1 Announce Type: cross Abstract: Real-world crisis intervention is inherently conversational, yet existing research largely focuses on static texts.Real-world crisis intervention is i

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Fact-Augmented Lookahead Planning for LLM Agents

DGX agent

arXiv:2506.09171v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly capable, but LLM agents still struggle to plan effectively in interactive, partially observable,

model-releasesarxiv-cs-ai
10 Jun 2026
Research

Human-AI Teaming Through the Lens of Calibration

DGX agent

arXiv:2606.10906v1 Announce Type: cross Abstract: We study models for human-AI teaming through the lens of statistical calibration. We assume the team consists of an AI model and human -- both of whic

researcharxiv-cs-ai
10 Jun 2026
Model Releases

Kwai Keye-VL-2.0 Technical Report

DGX agent

arXiv:2606.10651v1 Announce Type: new Abstract: We introduce Kwai Keye-VL-2.0-30B-A3B, an open-source Mixture-of-Experts (MoE) multimodal foundation model designed to advance long-video understanding

model-releasesarxiv-cs-cv
10 Jun 2026
Research

Linguistically Augmented Audio Speech Data (LinguAS)

DGX agent

arXiv:2606.10246v1 Announce Type: cross Abstract: Maliciously-created fake speech, including deepfaked and spoofed audio, is proliferating at an alarming rate, and detection models are racing to stay

researcharxiv-cs-ai
10 Jun 2026
Model Releases

LLM-as-a-Discriminator: When Synthetic Tables Still Look Real

DGX agent

arXiv:2606.09865v1 Announce Type: new Abstract: Privacy and data sharing are often in tension. Many organizations use synthetic data to reduce privacy risk and still share useful data. For tabular dat

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

Local Is Not a Sufficient Privacy Boundary: Governing OS-Integrated On-Device AI

DGX agent

arXiv:2606.10173v1 Announce Type: cross Abstract: As AI systems move into operating systems, privacy no longer turns only on whether a model runs locally. A local assistant may assemble email, calenda

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

MIRAGE: A Polarity-Flipping Encoding Subspace in LLM Agents

DGX agent

arXiv:2606.10304v1 Announce Type: new Abstract: When LLM agents are coerced into covertly encoding sensitive data (Base64, ROT13, acrostic, synonym chains, and beyond), the resulting outputs evade out

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

Optimal Post-Training Quantization Scales and Where to Find Them

DGX agent

arXiv:2606.10890v1 Announce Type: cross Abstract: Post-training quantization (PTQ) compresses large language models by mapping weights to low-bit representations. The scaling factor that defines the q

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Piper: A Programmable Distributed Training System

DGX agent

arXiv:2606.11169v1 Announce Type: cross Abstract: Large-scale model training increasingly relies on composing multiple parallelism strategies, such as data, pipeline, and expert parallelism, together

model-releasesarxiv-cs-ai
10 Jun 2026
Research

Pre-AF 13: An Interpretable Atrial Fibrillation Risk Score Mined from Discharge Reports

DGX agent

arXiv:2606.10725v1 Announce Type: cross Abstract: Background. Atrial fibrillation (AF) is the most prevalent cardiac arrhythmia and a major determinant of prognosis. Established AF risk scores rely on

researcharxiv-cs-cl
10 Jun 2026
Research

RankLLM: Weighted Ranking of LLMs by Quantifying Question Difficulty

DGX agent

arXiv:2602.12424v2 Announce Type: replace-cross Abstract: Benchmarks establish a standardized evaluation framework to systematically assess the performance of large language models (LLMs), facilitatin

researcharxiv-cs-ai
10 Jun 2026
Model Releases

SPACE: Source-free Proxy Anchor Concept Erasure for MLLMs

DGX agent

arXiv:2606.09868v1 Announce Type: cross Abstract: As Multimodal Large Language Models (MLLMs) face growing privacy risks and regulatory constraints, machine unlearning (MU) has emerged as a crucial so

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

The Interlocutor Effect: Why LLMs Leak More Personal Data to Agents Than Humans

DGX agent

arXiv:2606.09844v1 Announce Type: cross Abstract: Large Language Models (LLMs) alter their privacy behavior based on the perceived identity of their interlocutor. While safety mechanisms typically pre

model-releasesarxiv-cs-ai
10 Jun 2026
Research

VFUSE: Virulent Feature Understanding with Sparse autoEncoders

DGX agent

arXiv:2606.10080v1 Announce Type: cross Abstract: Generative models have shown remarkable progress in a variety of domains such as protein design, but such power enables the opaque generation of hazar

researcharxiv-cs-ai
10 Jun 2026
← Previous
1…314315316317318…1065
Next →