AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

Content type
88,271Total entries
1Added by human
88,270Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,935 results
Model Releases

Invariant Gradient Alignment for Robust Reasoning Distillation

DGX agent

arXiv:2606.05025v1 Announce Type: cross Abstract: Large language models (LLMs) suffer from shortcut learning: they systematically fail on out-of-distribution (OOD) inputs whose semantic surface differ

model-releasesarxiv-cs-ai
4 Jun 2026
Research

L^3: Large Lookup Layers

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2601.21461v3 Announce Type: replace-cross Abstract: Modern sparse language models typically achieve sparsity through Mixture-of-Experts (MoE) layers, which dynamically route tokens to dense MLP

researcharxiv-cs-ai
4 Jun 2026
Model Releases

MeshTok: Efficient Multi-Scale Tokenization for Scalable PDE Transformers

DGX agent

arXiv:2606.04366v1 Announce Type: new Abstract: Conventional patchified Transformers operate on uniform spatial partitions, distributing computational effort evenly across the domain irrespective of l

model-releasesarxiv-cs-lg
4 Jun 2026
Local Ai

MeshWeaver: Sparse-Voxel-Guided Surface Weaving for Autoregressive Mesh Generation

DGX agent

arXiv:2606.04688v1 Announce Type: new Abstract: Autoregressive mesh generation has gained attention by tokenizing meshes into sequences and training models in a language-modeling fashion. However, exi

local-aiarxiv-cs-cv
4 Jun 2026
Agents

MetaPoint: Unlocking Precise Spatial Control in Agentic Visual Generation

DGX agent

arXiv:2606.05031v1 Announce Type: new Abstract: Generative visual models fundamentally struggle with precise spatial control. This arises from a core disconnect: models can process textual description

agentsarxiv-cs-cv
4 Jun 2026
Safety

MusaCoder: Native GPU Kernel Generation with Full-Stack Training on Moore Threads GPU

DGX agent

arXiv:2606.04847v1 Announce Type: cross Abstract: Native GPU kernel generation turns high-level tensor programs into executable, efficient low-level code. Existing Large Language Models (LLMs) struggl

safetyarxiv-cs-cl
4 Jun 2026
Model Releases

PersistBench: When Should Long-Term Memories Be Forgotten by LLMs?

DGX agent

arXiv:2602.01146v2 Announce Type: replace Abstract: Conversational assistants are increasingly integrating long-term memory with large language models (LLMs). This persistence of memories, e.g., the u

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

RAVQ-HoloNet: Rate-Adaptive Vector-Quantized Hologram Compression

DGX agent

arXiv:2511.21035v2 Announce Type: replace Abstract: Holography offers significant potential for AR/VR applications. However, its adoption is limited by the high demand for data compression. Existing d

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

Shifting the Breaking Point of Flow Matching for Multi-Instance Editing

DGX agent

arXiv:2602.08749v3 Announce Type: replace Abstract: Flow matching models have recently emerged as an efficient alternative to diffusion, especially for text-guided image generation and editing, offeri

model-releasesarxiv-cs-cv
4 Jun 2026
Model Releases

Signed Dual Attention: Capturing Signed Dependencies in Time Series Forecasting

DGX agent

arXiv:2606.04833v1 Announce Type: cross Abstract: Initially developed for natural language processing, Transformer architectures and attention mechanisms are now central to a wide range of deep learni

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

SMADE-IE: Sparse Multi-Agent Framework with Evidence-Driven Debate for Zero-Shot Information Extraction

DGX agent

arXiv:2606.04691v1 Announce Type: new Abstract: Zero-shot information extraction (IE) with large language models (LLMs) has attracted increasing attention due to its flexibility in adapting to new sch

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

Solving Zebra Puzzles Using Constraint-Guided Multi-Agent Systems

DGX agent

arXiv:2407.03956v3 Announce Type: replace-cross Abstract: Prior research has enhanced the ability of Large Language Models (LLMs) to solve logic puzzles using techniques such as chain-of-thought promp

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

StandardE2E: A Unified Framework for End-to-End Autonomous Driving Datasets

DGX agent

arXiv:2606.04271v1 Announce Type: cross Abstract: Autonomous driving has shifted from modular perception-prediction-planning stacks toward end-to-end (E2E) models that map sensor inputs directly to ve

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Structure-Aware Prediction of PROTAC-Mediated Protein Degradability via Graph Neural Networks

DGX agent

arXiv:2606.04021v1 Announce Type: cross Abstract: Proteolysis-targeting chimeras (PROTACs) can selectively degrade disease-causing proteins, yet predicting which targets are amenable to degradation re

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

TaDA: Calibrated Probe Gating for Task-Domain LoRA Merging

DGX agent

arXiv:2606.05016v1 Announce Type: new Abstract: Combining a task LoRA adapter with a domain LoRA adapter into a single unified model is a practical yet largely unexplored challenge. Existing methods t

model-releasesarxiv-cs-cl
4 Jun 2026
Research

TANDEM: Bi-Level Data Mixture Optimization with Twin Networks

DGX agent

arXiv:2606.04401v1 Announce Type: new Abstract: The capabilities of large language models (LLMs) significantly depend on training data drawn from various domains. Optimizing domain-specific mixture ra

researcharxiv-cs-lg
4 Jun 2026
Model Releases

Thinking Through Signs: PEEL as a Semiotic Scaffolding for Epistemically Accountable AI-Enabled Research

DGX agent

arXiv:2606.04152v1 Announce Type: new Abstract: Large language models are reshaping research practice while quietly eroding researchers epistemic accountability. This commentary introduces PEEL - Prot

model-releasesarxiv-cs-ai
4 Jun 2026
Agents

Toward Autonomous O-RAN: A Multi-Scale Agentic AI Framework for Real-Time Network Control and Management

DGX agent

arXiv:2602.14117v2 Announce Type: replace-cross Abstract: Open Radio Access Networks (O-RAN) promise flexible 6G network access through disaggregated, software-driven components and open interfaces, b

agentsarxiv-cs-ai
4 Jun 2026
Model Releases

Toward Pre-Deployment Assurance for Enterprise AI Agents: Ontology-Grounded Simulation and Trust Certification

DGX agent

arXiv:2606.04037v1 Announce Type: new Abstract: Pre-deployment verification of enterprise artificial intelligence (AI) agents remains a critical gap between large language model (LLM) capability bench

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Towards Efficient and Evidence-grounded Mobility Prediction with LLM-Driven Agent

DGX agent

arXiv:2606.05130v1 Announce Type: cross Abstract: Individual-level mobility prediction is central to urban simulation, transportation planning, and policy analysis. Supervised sequence models achieve

model-releasesarxiv-cs-ai
4 Jun 2026
Agents

Trivium: Temporal Regret as a First-Class Objective for Causal-Memory Controllers

DGX agent

arXiv:2606.04421v1 Announce Type: new Abstract: Many current agentic systems and LLM pipelines correct mistakes by optimizing outcome reward. This addresses only the what of failure: when an outcome d

agentsarxiv-cs-ai
4 Jun 2026
Model Releases

Unpredictable Safety: Domain-Dependent Compliance and the Transparency Gap in Open-Weight LLMs

DGX agent

arXiv:2606.04035v1 Announce Type: cross Abstract: We present a systematic study of domain-dependent safety behavior in open-weight LLMs: 7 standardized experiments across 7 ethical domains, testing 5

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

What If Prompt Injection Never Left? Exploring Cross-Session Stored Prompt Injection in Agentic Systems

DGX agent

arXiv:2606.04425v1 Announce Type: cross Abstract: Modern agentic systems transform LLMs from session-bounded assistants into stateful systems that persist and evolve shared world state across sessions

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

When Seeing Is Not Believing -- A Benchmark for Search-Grounded Video Misinformation Detection

DGX agent

arXiv:2606.04098v1 Announce Type: new Abstract: Video misinformation increasingly operates at the semantic and evidential level: authentic footage may be selectively edited, temporally reordered, spli

model-releasesarxiv-cs-cv
4 Jun 2026
Model Releases

Analytical Evaluation of DCA Convergence Properties for Minimizing Prediction Functions of Gaussian RBF Support Vector Regression

DGX agent

arXiv:2606.03559v1 Announce Type: new Abstract: For nonconvex optimization problems whose objective is the prediction function of a trained Support Vector Regression (SVR) model with the Gaussian radi

model-releasesarxiv-cs-lg
3 Jun 2026
Model Releases

AnyAudio-Judge: A Dynamic Rubric-Based Benchmark and Evaluator for Audio Instruction Following

DGX agent

arXiv:2606.03116v1 Announce Type: cross Abstract: The rapid advancement of instruction-guided audio generation has highlighted the critical need for robust alignment evaluation. Current automated eval

model-releasesarxiv-cs-ai
3 Jun 2026
Safety

Attention Calibration for Position-Fair Dense Information Retrieval

DGX agent

arXiv:2606.02737v1 Announce Type: cross Abstract: Dense retrieval models exhibit positional bias: retrieval effectiveness degrades when relevant information appears later in a passage (Zeng et al., 20

safetyarxiv-cs-ai
3 Jun 2026
Model Releases

Beyond Ideal Instruction: A Comprehensive Framework for Evaluating LLMs in Realistic Interactions

DGX agent

arXiv:2606.03318v1 Announce Type: new Abstract: Despite great advances in tool-use capabilities of large language models (LLMs), existing evaluation benchmarks struggle to fully align with real-world

model-releasesarxiv-cs-cl
3 Jun 2026
Model Releases

Calibration Data Trade-offs Across Capability Dimensions: Why Multi-Source Mixing Matters for High-Sparsity LLM Pruning

DGX agent

arXiv:2606.03328v1 Announce Type: cross Abstract: Post-training pruning compresses large language models to high sparsity using a small unlabelled calibration set, and recent work has concluded that t

model-releasesarxiv-cs-ai
3 Jun 2026
Safety

Consistency Training Can Entrench Misalignment

DGX agent

arXiv:2606.03810v1 Announce Type: cross Abstract: Consistency training encourages a model to produce similar outputs across related inputs or sampling procedures. Such methods are simple, scalable, an

safetyarxiv-cs-ai
3 Jun 2026
Local Ai

Demystifying Pipeline Parallelism: First Theory for PipeDream

DGX agent

arXiv:2606.03498v1 Announce Type: new Abstract: Training modern machine learning models increasingly requires computation to be distributed across many accelerators. Data parallelism remains the defau

local-aiarxiv-cs-lg
3 Jun 2026
Model Releases

Diagnosis of Human Object Interaction Detectors for Real World Educational Applications

DGX agent

arXiv:2606.02789v1 Announce Type: new Abstract: Human-object interaction (HOI) recognition is critical for automatically analyzing student behavior in complex educational environments. Although state-

model-releasesarxiv-cs-cv
3 Jun 2026
Model Releases

Evaluating and Calibrating LLM Confidence on Questions with Multiple Correct Answers

DGX agent

arXiv:2602.07842v2 Announce Type: replace Abstract: Confidence calibration is essential for making large language models (LLMs) reliable, yet existing training-free methods have been primarily studied

model-releasesarxiv-cs-cl
3 Jun 2026
Research

Fast and Expressive Multi-Byte Prediction with Probabilistic Circuits

DGX agent

arXiv:2511.11346v2 Announce Type: replace Abstract: Multi-token prediction (MTP) is a prominent strategy to significantly speed up generation in large language models (LLMs), especially in byte-level

researcharxiv-cs-lg
3 Jun 2026
Safety

Fast Organic Crystal Structure Prediction with Unit Cell Flow Matching

DGX agent

arXiv:2606.03199v1 Announce Type: new Abstract: Organic crystal structure prediction (CSP) is a requirement for computational modelling of organic solids, but traditionally costs several CPU-years per

safetyarxiv-cs-lg
3 Jun 2026
Model Releases

FlashMLA-ETAP: Efficient Transpose Attention Pipeline for Accelerating MLA Inference on NVIDIA H20 GPUs

DGX agent

arXiv:2506.01969v3 Announce Type: replace-cross Abstract: Efficient inference of Multi-Head Latent Attention (MLA) is challenged by deploying the DeepSeek-R1 671B model on a single Multi-GPU server. T

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

From Long News to Accurate Forecast: Importance-Aware Fusion and PRM-Guided Reflection for Time Series Forecasting

DGX agent

arXiv:2606.03097v1 Announce Type: new Abstract: Incorporating news into time series forecasting is appealing because news can reveal abrupt exogenous events that historical values alone cannot recover

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Gate AI: LLM Security Benchmark Evaluation Methodology and Results

DGX agent

arXiv:2606.02959v1 Announce Type: new Abstract: Published evaluations of prompt-injection and jailbreak detectors for Large Language Models often suffer from two systematic weaknesses: per-dataset thr

model-releasesarxiv-cs-lg
3 Jun 2026
Model Releases

Geometry-Aware Tabular Diffusion

DGX agent

arXiv:2606.02607v1 Announce Type: cross Abstract: Tabular synthesis is critical for privacy-preserving sharing and augmentation, yet diffusion models rely on implicit mechanisms to capture inter-colum

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Hallucinations as Orthogonal Noise: Inference-Time Manifold Alignment via Dynamic Contextual Orthogonalization

DGX agent

arXiv:2606.03022v1 Announce Type: cross Abstract: Hallucination in Large Language Models (LLMs), characterized by the generation of content inconsistent with contextual facts or logical constraints --

model-releasesarxiv-cs-ai
3 Jun 2026
Tutorials

HiSE: A Lightweight Hierarchical Semantic Explainer for Heterogeneous Graph Neural Networks

DGX agent

arXiv:2606.03495v1 Announce Type: new Abstract: Heterogeneous graph neural networks (HGNNs) have demonstrated remarkable performance in modeling complex relational data, however their interpretability

tutorialsarxiv-cs-lg
3 Jun 2026
Model Releases

HyperPatch: Sequential Knowledge Editing Under n-ary Structural Drift

DGX agent

arXiv:2606.03179v1 Announce Type: new Abstract: Large Language Models (LLMs) rely on Knowledge Editing (KE) to maintain temporal validity, yet real-world knowledge is inherently n-ary. We demonstrate

model-releasesarxiv-cs-cl
3 Jun 2026
Local Ai

Learn from Your Mistakes: Tree-like Self-Play for Secure Code LLMs

DGX agent

arXiv:2606.03489v1 Announce Type: cross Abstract: While Large Language Models (LLMs) excel in code generation, they remain prone to replicating subtle yet critical vulnerabilities endemic to their tra

local-aiarxiv-cs-ai
3 Jun 2026
Safety

Learning Self-Interpretation from Interpretability Artifacts: Training Lightweight Adapters on Vector-Label Pairs

DGX agent

arXiv:2602.10352v2 Announce Type: replace-cross Abstract: Self-interpretation methods prompt language models to describe their own internal states, but remain unreliable due to hyperparameter sensitiv

safetyarxiv-cs-ai
3 Jun 2026
Model Releases

MemoGen: Can Past Experience Improve Future Text-to-Image Generation?

DGX agent

arXiv:2606.03243v1 Announce Type: new Abstract: Modern text-to-image models have achieved strong visual synthesis, yet remain unreliable when prompts require implicit visual constraints, relational re

model-releasesarxiv-cs-cv
3 Jun 2026
Model Releases

Message Tuning Outshines Graph Prompt Tuning: A Prismatic Space Perspective

DGX agent

arXiv:2606.03290v1 Announce Type: cross Abstract: Graph Foundation Models (GFMs), built upon the Pre-training and Adaptation paradigm, have emerged as a research hotspot in graph learning. For GNN-bas

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Multi^2: Hierarchical Multi-Agent Decision-Making with LLM-Based Agents in Interactive Environments

DGX agent

arXiv:2606.03698v1 Announce Type: new Abstract: A central goal of large language model (LLM) research is to build agentic systems that can plan, act, and adapt through sustained interaction with dynam

model-releasesarxiv-cs-lg
3 Jun 2026
Model Releases

NeuroArmor: Safe-Variant-Guided Representation Consistency for Selective Re-Anchoring in Jailbreak Defense

DGX agent

arXiv:2606.03486v1 Announce Type: cross Abstract: Large language models remain vulnerable to jailbreak attacks that hide harmful intent behind seemingly ordinary requests such as role-play, translatio

model-releasesarxiv-cs-ai
3 Jun 2026
← Previous
1…417418419420421…1082
Next →