AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlog
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,794 results
Model Releases

Pragmatic Reasoning improves LLM Code Generation

DGX agent

arXiv:2502.15835v5 Announce Type: replace-cross Abstract: Pragmatic reasoning helps interlocutors infer intended meaning from ambiguous or underspecified messages by considering shared context and cou

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

QUEST: Training Frontier Deep Research Agents with Fully Synthetic Tasks

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2605.24218v1 Announce Type: new Abstract: Deep research agents extend the role of search engines from retrieving keyword-matched pages to synthesizing knowledge, fundamentally changing how human

model-releasesarxiv-cs-cl
26 May 2026
Safety

Reading, Not Thinking: Understanding and Bridging the Modality Gap When Text Becomes Pixels in Multimodal LLMs

DGX agent

arXiv:2603.09095v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) can process text presented as images, yet they often perform worse than when the same content is provided a

safetyarxiv-cs-cl
26 May 2026
Model Releases

Residual Drift Dominates Contradiction in Multi-Turn Constraint Reasoning

DGX agent

arXiv:2605.23940v1 Announce Type: new Abstract: How do multi-turn reasoning systems fail? The expected answer is logical contradiction, in which the system's maintained state becomes unsatisfiable. We

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

RoboManipBaselines: A Unified Framework for Imitation Learning in Robotic Manipulation across Real and Simulation Environments

DGX agent

arXiv:2509.17057v3 Announce Type: replace Abstract: We present RoboManipBaselines, an open-source software framework for imitation learning research in robotic manipulation. The framework supports the

model-releasesarxiv-cs-ro
26 May 2026
Model Releases

Robust Fuzzy Multi-view Learning under View Conflict

DGX agent

arXiv:2605.24475v1 Announce Type: cross Abstract: Trusted multi-view classification aims to deliver reliable fusion for accurate predictions and has recently attracted substantial attention in both ac

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Spectral Probe-Circuits: A Three-Step Recipe for Identifying Attention-Head Circuits in Pretrained Transformers

DGX agent

arXiv:2605.24059v1 Announce Type: cross Abstract: We present a three-step recipe for identifying attention-head circuits in pretrained transformers. A per-head spectral signal -- the time-integrated p

model-releasesarxiv-cs-ai
26 May 2026
Safety

TapSampling: Inference-Time Sampling with a Task-Progress-Understanding Verifier for Robotic Manipulation

DGX agent

arXiv:2605.25547v1 Announce Type: new Abstract: Existing embodied control research demonstrates remarkable performance improvements by scaling training data and model size. We instead explore inferenc

safetyarxiv-cs-ro
26 May 2026
Safety

The Concept Allocation Zone: Tracking How Concepts Form Across Transformer Depth

DGX agent

arXiv:2605.24856v1 Announce Type: cross Abstract: Concept formation in transformer language models is depth-extended, not a single-layer event: concepts emerge gradually across a contiguous region of

safetyarxiv-cs-ai
26 May 2026
Model Releases

Uni-DPO: A Unified Paradigm for Dynamic Preference Optimization of LLMs

DGX agent

arXiv:2506.10054v4 Announce Type: replace-cross Abstract: Direct Preference Optimization (DPO) has emerged as a cornerstone of reinforcement learning from human feedback (RLHF) due to its simplicity a

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

v0.13.0 of Exo just dropped, and it's one I've been looking forward to for a while. tl;dr swapping out anthropic for @ollama cloud dropped m…

DGX agent

v0.13.0 of Exo just dropped, and it's one I've been looking forward to for a while. tl;dr swapping out anthropic for @ollama cloud dropped my usage costs from 30/day to just 20 per MONTH with no notic

model-releasesollama--x
26 May 2026
Applications

VectorArk: Learning Practical Image Vectorization with Rounded Polygon Representation

DGX agent

arXiv:2605.24398v1 Announce Type: cross Abstract: Recent vision-language model (VLM)-based approaches have achieved impressive results on image vectorization tasks. However, they are typically evaluat

applicationsarxiv-cs-ai
26 May 2026
Model Releases

WhoSaidIt: Human-LLM Collaborative Annotation for Text-Based Multilingual Speaker-Attribute Classification

DGX agent

arXiv:2605.26070v1 Announce Type: new Abstract: Annotating speaker attributes from text is inherently ambiguous, particularly in multilingual settings where demographic and social cues are implicit an

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

AI Evaluation Should Require Standardized Item-Level Data Releases

DGX agent

arXiv:2604.03244v2 Announce Type: replace Abstract: This position paper argues that standardized item-level benchmark data should become the default infrastructure for AI evaluation. Current evaluatio

model-releasesarxiv-cs-ai
25 May 2026
Safety

ALIVE: Awakening LLM Reasoning via Adversarial Learning and Instructive Verbal Evaluation

DGX agent

arXiv:2602.05472v2 Announce Type: replace Abstract: The quest for expert-level reasoning in Large Language Models (LLMs) has been hampered by a persistent extit{reward bottleneck}: traditional reinfor

safetyarxiv-cs-ai
25 May 2026
Research

Any-Dimensional Invariant Universality

DGX agent

arXiv:2605.23156v1 Announce Type: new Abstract: Several machine learning models are defined for inputs of any size, such as graphs with different numbers of nodes and point clouds containing varying n

researcharxiv-cs-lg
25 May 2026
Safety

BarrierSteer: LLM Safety via Learning Barrier Steering

DGX agent

arXiv:2602.20102v2 Announce Type: replace-cross Abstract: Despite the strong performance of large language models (LLMs) across diverse tasks, their susceptibility to adversarial attacks and unsafe co

safetyarxiv-cs-ai
25 May 2026
Safety

Contrastive Distribution Matching for Amortized Sequential Monte Carlo in Discrete Diffusion

DGX agent

arXiv:2605.23346v1 Announce Type: new Abstract: Discrete diffusion models have emerged as powerful frameworks for generating structured categorical data. However, efficiently sampling from reward-tilt

safetyarxiv-cs-lg
25 May 2026
Model Releases

DDX-TRACE: A Benchmark for Medical Diagnostic Trajectories in VLMs

DGX agent

arXiv:2605.23629v1 Announce Type: new Abstract: Medical diagnosis is not a single prediction from a fully specified vignette. It is a sequential workup: clinicians decide what evidence to obtain, revi

model-releasesarxiv-cs-cv
25 May 2026
Model Releases

Decoupling Spatio-Temporal Adapter for Fine-Grained Badminton Action Localization

DGX agent

arXiv:2605.23355v1 Announce Type: new Abstract: Temporal Action Localization (TAL) has been extensively studied in generic video understanding, while fine-grained sports scenarios, such as professiona

model-releasesarxiv-cs-cv
25 May 2026
Research

Enhancing Blood Cells Classification using Hybrid Quantum Neural Networks

DGX agent

arXiv:2605.23324v1 Announce Type: new Abstract: Accurate classification of microscopic blood cells is still a critical task in medical image analysis, where subtle variations and limited data can chal

researcharxiv-cs-cv
25 May 2026
Industry

It’ll be interesting to see if the post training for this uses a multiple of the compute of pretraining as cursor did when they tuned Kimi a…

DGX agent

It’ll be interesting to see if the post training for this uses a multiple of the compute of pretraining as cursor did when they tuned Kimi as the base model Grok foundation model V9-Medium (1.5T) has

industryemad-mostaque--x
25 May 2026
Model Releases

Lipschitz Optimization for Formal Verification of Homographies

DGX agent

arXiv:2605.23203v1 Announce Type: cross Abstract: The adoption of vision neural networks in regulated industries requires formal robustness guarantees, especially in safety-critical domains such as he

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

MedExpMem: Adapting Experience Memory for Differential Diagnosis

DGX agent

arXiv:2605.22872v1 Announce Type: cross Abstract: Experienced physicians develop diagnostic expertise through clinical practice, acquiring not only disease knowledge but also the ability to differenti

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

Pointwise Metrics Mislead: An Evaluation Protocol for Multimodal Inverse Problems

DGX agent

arXiv:2605.22891v1 Announce Type: new Abstract: Evaluation in scientific reconstruction is dominated by pointwise metrics - RMSE, MAE, per-event resolution - under the implicit assumption that lower e

model-releasesarxiv-cs-lg
25 May 2026
Model Releases

Reinforcement Learning for Microcanonical Graph Ensemble with Assortativity Constraints

DGX agent

arXiv:2605.23285v1 Announce Type: cross Abstract: How network structure determines function is a fundamental question, and it can be investigated by graph ensembles with precisely controlled structura

model-releasesarxiv-cs-ai
25 May 2026
Research

Robust LLM Watermarking with Minimal Semantic Distortion for IP Protection

DGX agent

arXiv:2605.23175v1 Announce Type: cross Abstract: Proprietary large language models (LLMs) face risks of intellectual property (IP) violation, as adversaries can replicate an LLM by collecting input-o

researcharxiv-cs-cl
25 May 2026
Model Releases

Semantically Structured Mixture-of-Experts for Compositional Robotic Manipulation

DGX agent

arXiv:2605.23477v1 Announce Type: new Abstract: Diffusion-based policies have established a new standard for precise robotic manipulation but face a critical scalability bottleneck: high-performance m

model-releasesarxiv-cs-ro
25 May 2026
Model Releases

SemEval-2026 Task 6: CLARITY -- Unmasking Political Question Evasions

DGX agent

arXiv:2603.14027v2 Announce Type: replace Abstract: Political speakers often avoid answering questions directly while maintaining the appearance of responsiveness. Despite its importance for public di

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

STAMBRIDGE: Spectral-Temporal Amplitude-aware Mid-Feature Bridge for EEG Visual Decoding

DGX agent

arXiv:2605.23137v1 Announce Type: cross Abstract: Electroencephalography (EEG) visual decoding remains challenging due to the modality gap between low-SNR neural signals and highly structured vision--

model-releasesarxiv-cs-cv
25 May 2026
Model Releases

Targeted Regularization for Causal Effect Estimation with Exponential Dispersion Family Outcomes

DGX agent

arXiv:2502.07295v2 Announce Type: replace Abstract: Neural Networks (NNs) for causal effect estimation have shown strong empirical performance, yet endowing them with desirable semiparametric properti

model-releasesarxiv-cs-lg
25 May 2026
Tutorials

The TIME Machine: On The Power of Motion for Efficient Perception

DGX agent

arXiv:2605.23045v1 Announce Type: cross Abstract: Video representation learning has seen tremendous progress in recent years. This has been driven by many factors, including the scale of training and

tutorialsarxiv-cs-ai
25 May 2026
Model Releases

Using Ensemble Diffusion to Estimate Uncertainty for End-to-End Autonomous Driving

DGX agent

arXiv:2506.00560v2 Announce Type: replace-cross Abstract: End-to-end planning systems for autonomous driving are rapidly improving, especially in closed-loop simulation environments like CARLA. Many s

model-releasesarxiv-cs-cv
25 May 2026
Model Releases

VideoTemp-o3: Harmonizing Temporal Grounding and Video Understanding in Agentic Thinking-with-Videos

DGX agent

arXiv:2602.07801v4 Announce Type: replace-cross Abstract: In long-video understanding, conventional uniform frame sampling often fails to capture key visual evidence, leading to degraded performance a

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

VINS-120K: Ultra High-Resolution Image Editing with A Large-Scale Dataset

DGX agent

arXiv:2605.23518v1 Announce Type: new Abstract: Directly editing ultra-high-resolution (UHR) images is valuable but underexplored, primarily due to the lack of high-quality data and the challenge in m

model-releasesarxiv-cs-cv
25 May 2026
Agents

When Is Next-Token Prediction Useful? Marginalization, Ergodicity, Mixture Identifiability, Local Sufficiency, RAG, Tools, and Programming

DGX agent

arXiv:2605.23278v1 Announce Type: new Abstract: Language models trained on observed sequences are often described as learning the conditional distribution of the next token given previous tokens. This

agentsarxiv-cs-cl
25 May 2026
Local Ai

Dual 3090s?

DGX agent

A discussion about difficulties getting Ollama to fully utilize dual NVIDIA RTX 3090 GPUs when running various language models including 8B and 70B parameter models. Users report that despite Ollama r

local-air-ollama
24 May 2026
Model Releases

A Boundary-Layer Mechanism for One-Third Scaling in Online Softmax Classification

DGX agent

arXiv:2605.22341v1 Announce Type: new Abstract: Hard-label classification is usually trained with smooth surrogate losses, most prominently softmax cross-entropy. We isolate an asymptotic mechanism by

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Alike Parts: A Feature-Informed Approach to Local and Global Prototype Explanations

DGX agent

arXiv:2605.21646v1 Announce Type: new Abstract: Prototype-based explanations offer an intuitive, example-based approach to support the interpretability of machine learning black box classifiers but of

model-releasesarxiv-cs-lg
23 May 2026
Local Ai

AsymFLUX.2-klein-9B is all about textures

DGX agent

AsymFLUX.2-klein-9B is a pixel-space text-to-image model finetuned from FLUX.2-klein-base-9B using the AsymFlow method. The model is noted for producing sharp textures with strong adherence and compos

local-air-stablediffusion
23 May 2026
Research

Asymmetric Virtual Memory Paging for Hybrid Mamba-Transformer Inference

DGX agent

arXiv:2605.22416v1 Announce Type: new Abstract: Hybrid language models like Jamba mix attention layers with State Space Models (SSMs), creating two memory cache types with opposite profiles: Key-Value

researcharxiv-cs-lg
23 May 2026
Research

Flashlight: PyTorch Compiler Extensions to Accelerate Attention Variants

DGX agent

arXiv:2511.02043v4 Announce Type: replace Abstract: Attention is a fundamental building block of large language models (LLMs), so there have been many efforts to implement it efficiently. For example,

researcharxiv-cs-lg
23 May 2026
Model Releases

Holomorphic Neural ODEs with Kolmogorov-Arnold Networks for Interpretable Discovery of Complex Dynamics

DGX agent

arXiv:2605.22235v1 Announce Type: new Abstract: Complex dynamical systems governed by holomorphic maps such as z^2 + c exhibit fractal boundaries with extreme sensitivity to initial conditions. Accura

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Reasoning through Verifiable Forecast Actions: Consistency-Grounded RL for Financial LLMs

DGX agent

arXiv:2605.21975v1 Announce Type: new Abstract: Financial markets are characterized by extreme non-stationarity, low signal-to-noise ratios, and strong dependence on external information such as news,

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

Rule-State Inference (RSI): A Bayesian Framework for Compliance Monitoring in Rule-Governed Domains

DGX agent

arXiv:2603.21610v2 Announce Type: replace Abstract: Compliance monitoring in rule-governed domains (tax administration, clinical protocol adherence, environmental regulation) faces three structural ob

model-releasesarxiv-cs-lg
23 May 2026
Research

Same Architecture, Different Capacity: Optimizer-Induced Spectral Scaling Laws

DGX agent

arXiv:2605.21803v1 Announce Type: new Abstract: Scaling laws have made language-model performance predictable from model size, data, and compute, but they typically treat the optimizer as a fixed trai

researcharxiv-cs-lg
23 May 2026
Model Releases

Self-Supervised ConvLSTM for Fermi Large Area Telescope Transient Detection

DGX agent

arXiv:2605.22112v1 Announce Type: cross Abstract: We present a framework for detecting transient gamma-ray phenomena in a controlled environment by combining end-to-end simulations of the Fermi-LAT sk

model-releasesarxiv-cs-lg
23 May 2026
Agents

Skill Weaving: Efficient LLM Improvement via Modular Skillpacks

DGX agent

arXiv:2605.22205v1 Announce Type: cross Abstract: Large language models increasingly require specialization across diverse domains, yet existing approaches struggle to balance multi-domain capacities

agentsarxiv-cs-lg
23 May 2026
← Previous
1…609610611612613…1371
Next →