AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

Content type
87,042Total entries
1Added by human
87,041Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,106 results
Model Releases

Why Are GUI Agents Correct but Late? Decode on the Decision-Time Critical Path, Tested with Pre-Compiled Policy Trees

DGX agent

arXiv:2607.28399v1 Announce Type: new Abstract: Computer-use agents often fail on transient GUI events because they produce the correct action only after the relevant window has already closed. We ide

model-releasesarxiv-cs-lg
31 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Aligning LLM-Simulated and Human Examinees for Psychometric Calibration: A Cognitive Diagnostic Profiling Approach

DGX agent

arXiv:2607.26317v1 Announce Type: cross Abstract: Psychometric calibration for educational tests typically requires costly human response data. Large language models (LLMs) simulated examinees offer a

model-releasesarxiv-cs-cl
30 Jul 2026
Applications

ContactFlow: A video action conditioning that transfers across embodiments

DGX agent

arXiv:2607.26579v1 Announce Type: cross Abstract: World models offer a promising route toward robot planning by enabling agents to imagine and verify the consequences of actions before execution. Howe

applicationsarxiv-cs-cv
30 Jul 2026
Model Releases

Crossing-Free Probabilistic K-Line Forecasts Without Retraining

DGX agent

arXiv:2607.26792v1 Announce Type: cross Abstract: Probabilistic K-line forecasting describes uncertainty in four complementary prices, namely open--high--low--close (OHLC). However, it introduces two

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Diagnosing Fine-Grained Inconsistency Classification in Financial Disclosure Text

DGX agent

arXiv:2607.26368v1 Announce Type: new Abstract: Financial disclosures contain numerical claims, temporal statements, entity references, policy commitments, and risk descriptions that may conflict in q

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

EgoSafe: A First-Person Mobile-Captured Benchmark for Visual Safety Understanding

DGX agent

arXiv:2607.26518v1 Announce Type: new Abstract: Reliable visual safety understanding in real-world scenarios demands more than just object recognition; it requires causal reasoning under epistemic unc

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

Enhancing Automated Machine Learning via Homogeneous Train-Test Splitting Methods

DGX agent

arXiv:2607.26625v1 Announce Type: new Abstract: Accurate model evaluation in machine learning depends critically on how datasets are split into training and testing subsets. Standard random splitting

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

FedWeave: Rethinking the Unit of Specialization in Heterogeneous Federated MoE-LoRA

DGX agent

arXiv:2607.26618v1 Announce Type: cross Abstract: Federated PEFT enables LLMs to collaboratively adapt to decentralized private data without sharing raw examples. However, task heterogeneity across cl

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Financial Volatility and Risk Forecasting Incorporating a Larger Number of Realized Measures

DGX agent

arXiv:2411.17136v2 Announce Type: replace-cross Abstract: Realised volatility has become increasingly prominent in volatility forecasting due to its ability to capture intraday price fluctuations. Wit

model-releasesarxiv-cs-lg
30 Jul 2026
Applications

Learning Dynamic User Personas from Implicit Interaction Streams via Iterative Refinement

DGX agent

arXiv:2607.26473v1 Announce Type: cross Abstract: Personalizing large language models (LLMs) to individual users is essential for improving user experience, yet existing approaches typically rely on e

applicationsarxiv-cs-cl
30 Jul 2026
Model Releases

LLMET: Enabling Cross-Layer Evaluation of Emerging M3D Memories for Energy-Efficient LLM Serving

DGX agent

arXiv:2607.26491v1 Announce Type: cross Abstract: The energy consumption of Large Language Model (LLM) serving is becoming a major system challenge as deployment scales, driven by hardware power and t

model-releasesarxiv-cs-lg
30 Jul 2026
Agents

Mitigating Compounding Error via Video Representation Regularization

DGX agent

arXiv:2607.27036v1 Announce Type: new Abstract: Video diffusion-based world models enable long autoregressive video generation for robotics, autonomous driving and simulation tasks, yet sliding-window

agentsarxiv-cs-cv
30 Jul 2026
Research

Prior Directions: Why GUI Grounding Gets Locked in the Past

DGX agent

arXiv:2607.26913v1 Announce Type: new Abstract: Vision-language models often use descriptions of earlier visual states to make decisions about the current scene. When the scene changes, stale language

researcharxiv-cs-cv
30 Jul 2026
Model Releases

Temporally Centered SIGReg Improves Multi-Task LeWorldModel Learning: From Analysis to Method

DGX agent

arXiv:2607.26924v1 Announce Type: new Abstract: Recent work on LeWorldModel (LeWM) has shown that the Sketched Isotropic Gaussian Regularizer (SIGReg) enables stable end-to-end world-model learning fr

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Transformers Can Learn Rules They've Never Seen: Proof of Computation Beyond Interpolation

DGX agent

arXiv:2603.17019v2 Announce Type: replace Abstract: A central question in the debate over large language models is whether transformers can learn rules they have never seen, or whether they can only i

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Addressable Recall Compaction for Long Context-Window Control in AI Agents

DGX agent

arXiv:2607.25066v1 Announce Type: new Abstract: Long-horizon LLM agents accumulate reasoning traces, actions, and tool observations that can eventually exceed a model's fixed context window. Existing

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Agentic AI for Scientific Reasoning in Autonomous Quantum Sensing Experiments

DGX agent

arXiv:2607.25145v1 Announce Type: cross Abstract: We implement an agentic AI workflow built around a large language model (LLM) agent for autonomous experiments with nitrogen-vacancy (NV) centers in d

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

AI's Capability in Assisting Scientific Research in Physics, Astrophysics, and Cosmology II: Project Planning and Proposal Evaluation

DGX agent

arXiv:2607.25881v1 Announce Type: new Abstract: We investigate how well large language models (LLMs) can assist scientific project planning and proposal evaluation. One-page project plans were indepen

model-releasesarxiv-cs-cl
29 Jul 2026
Model Releases

At-the-Roofline Sparse Tensor Contractions on Vector Processors for Transformer Inference

DGX agent

arXiv:2607.25504v1 Announce Type: cross Abstract: Fine-grained weight pruning and activation sparsification have emerged as effective approaches for reducing the compute and memory cost of inference f

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Authoring Agent Skills: A Software-Engineering Approach

DGX agent

arXiv:2607.25032v1 Announce Type: cross Abstract: Agent Skills are an emerging way to extend large language model agents with reusable procedural knowledge that the agent loads on demand. Anthropic in

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Bridging Compute- and Data-Optimal Pretraining

DGX agent

arXiv:2607.25271v1 Announce Type: cross Abstract: Classical compute-optimal scaling laws assume an unbounded supply of fresh pretraining data, yet pretraining is increasingly entering a regime in whic

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

CARE-MH: Towards Unified, Reproducible, and Comparable Evaluation of Mental Health LLMs

DGX agent

arXiv:2607.24754v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to provide mental health support, requiring reliable evaluation of safety, empathy, and therapeutic

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

CLBench-V: Evaluating Multimodal Context Learning from Grounding to Knowledge Acquisition

DGX agent

arXiv:2607.25294v1 Announce Type: cross Abstract: Real-world tasks often require models to learn from task-specific context rather than relying only on pre-trained knowledge. While recent work has hig

model-releasesarxiv-cs-ai
29 Jul 2026
Safety

CoRT: Counterfactual Replay for Token-Level Rubric-Guided Policy Optimization

DGX agent

arXiv:2607.25659v1 Announce Type: new Abstract: Rubric-based reinforcement learning enriches language model training by evaluating model outputs against explicit criteria. Yet in GRPO-style pipelines,

safetyarxiv-cs-ai
29 Jul 2026
Model Releases

COVENANT: Natural-Language Workflow Compilation for Aligned Agent Execution

DGX agent

arXiv:2607.25400v1 Announce Type: new Abstract: Large language model (LLM) agents are increasingly entrusted with natural-language workflow instructions (e.g., retail-payment policies) that specify no

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Evaluating VLMs for Autonomous Agent-Driven Geometry Clipping Detection in Video Game QA

DGX agent

arXiv:2607.25921v1 Announce Type: cross Abstract: In this work, we study the use of Vision-Language Models (VLMs) for anomaly detection in an agent-driven game Quality Assurance (QA) pipeline focusing

model-releasesarxiv-cs-ai
29 Jul 2026
Research

From Training to Deployment: Post-Hoc Causal Feature Identification via Sensitivity Ratios

DGX agent

arXiv:2607.25546v1 Announce Type: new Abstract: Given a model that is already trained, which features does it rely on causally versus spuriously? Existing methods require access to the training proced

researcharxiv-cs-ai
29 Jul 2026
Research

Generative Distributionally Robust Optimization

DGX agent

arXiv:2607.24983v1 Announce Type: cross Abstract: Generative models are increasingly adopted in distributionally robust optimization (DRO), but existing approaches trade off model compatibility and ad

researcharxiv-cs-ai
29 Jul 2026
Model Releases

IMPRINT: Image-Conditioned Query Enrichment for Long-Tail Object Goal Navigation

DGX agent

arXiv:2607.25106v1 Announce Type: new Abstract: Embodied AI increasingly relies on queryable semantic maps built from pre-trained vision-language models to enable zero-shot Object Goal Navigation (Obj

model-releasesarxiv-cs-cv
29 Jul 2026
Model Releases

Med-R^3: Enhancing Medical Retrieval-Augmented Reasoning of LLMs via Progressive Reinforcement Learning

DGX agent

arXiv:2507.23541v5 Announce Type: replace Abstract: In medical scenarios, effectively retrieving external knowledge and leveraging it for rigorous logical reasoning is of significant importance. Despi

model-releasesarxiv-cs-cl
29 Jul 2026
Model Releases

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing

DGX agent

arXiv:2607.25300v1 Announce Type: new Abstract: Video editing is fundamentally message-driven: even from the same source footage, the selected shots change depending on the narrative the editor wishes

model-releasesarxiv-cs-cv
29 Jul 2026
Model Releases

OmniPhys: Knowledge-Graph-Driven Benchmarking and Collective Optimization for Physical Commonsense in Text-to-Image Generation

DGX agent

arXiv:2607.25641v1 Announce Type: cross Abstract: While text-to-image models exhibit remarkable visual fidelity, they frequently violate fundamental physical commonsense. Existing benchmarks often rel

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Reinforcement Learning for Code Optimization

DGX agent

arXiv:2607.25970v1 Announce Type: cross Abstract: RL for code correctness is now established: have the model generate a program, run it against hidden test cases, and reward solutions that pass. Exten

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Rethinking CD: A Reproducibility Study and Extension on the Ineffectiveness of Contrastive Decoding at Mitigating Object Hallucinations in MLLMs

DGX agent

arXiv:2607.25196v1 Announce Type: new Abstract: Contrastive decoding (CD) has been proposed as a training-free strategy for mitigating object hallucinations in multimodal large language models (MLLMs)

model-releasesarxiv-cs-lg
29 Jul 2026
Model Releases

ScaleResfusion: Residual Rectified Flow based on Residual Vector Field

DGX agent

arXiv:2607.25275v1 Announce Type: cross Abstract: Real-world Image Restoration (Real-IR) aims to recover high-quality (HQ) images from complex and unknown degradations. Although recent diffusion-based

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Agent Team Work Zone: An Automated, Persistent Workspace for Long-Lived Coding Agent Teams

DGX agent

arXiv:2607.22917v1 Announce Type: new Abstract: Large Language Model (LLM) agents have significantly improved coding and programming workflows. Claude Code, in particular, is one of the most powerful

model-releasesarxiv-cs-ai
28 Jul 2026
Research

Are Prompt Optimizers Blind? Cross-Modal Visual Feedback for Automatic Prompt Optimization

DGX agent

arXiv:2607.24354v1 Announce Type: new Abstract: Automatic prompt optimization (APO) has been widely adopted to adapt vision-language models (VLMs) to downstream tasks without weight updates, yielding

researcharxiv-cs-ai
28 Jul 2026
Model Releases

AutoMat: Enabling Automated Crystal Structure Reconstruction from Microscopy via Agentic Tool Use

DGX agent

arXiv:2505.12650v2 Announce Type: replace-cross Abstract: Reconstructing atomistic crystal structures from a single noisy STEM projection is an ill-posed inverse problem: multiple lattices can explain

model-releasesarxiv-cs-ai
28 Jul 2026
Safety

Breaking the Synthetic-Real Domain Shortcut for Training-Free Generative Replay-based Class Incremental Learning

DGX agent

arXiv:2607.22994v1 Announce Type: new Abstract: Class-incremental learning (CIL) requires models to continuously acquire new knowledge while avoiding catastrophic forgetting. While exemplar replay is

safetyarxiv-cs-cv
28 Jul 2026
Model Releases

CameraAnything: Refilming Videos with Arbitrary Camera Control

DGX agent

arXiv:2607.24591v1 Announce Type: new Abstract: We introduce CameraAnything, the first unified framework for camera controlled video editing that enables joint control of both intrinsic and extrinsic

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

Decentralized Granular Access Control for Agentic AI Systems in Critical Infrastructure

DGX agent

arXiv:2607.22611v1 Announce Type: new Abstract: The deployment of autonomous AI agents in production infrastructure introduces fundamental security challenges that traditional role-based access contro

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

DeepLook: Deeper Thinking with Lookahead

DGX agent

arXiv:2607.22602v1 Announce Type: new Abstract: Inference-time scaling has emerged as a powerful paradigm for improving large language model reasoning, often delivering larger gains on difficult reaso

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Diffusion-Guided Search via Exponential Tilting (DiffTilt): An Application to Falsification of Safety-Critical Systems

DGX agent

arXiv:2607.23134v1 Announce Type: new Abstract: Discovering rare safety-critical failures in autonomous and cyber-physical systems is a fundamental challenge in verification and validation. Existing f

model-releasesarxiv-cs-lg
28 Jul 2026
Model Releases

DraftExpert: Expansion-Aware Self-Speculative Decoding for End-Device MoE Inference

DGX agent

arXiv:2607.24434v1 Announce Type: cross Abstract: Large Mixture-of-Experts (MoE) language models are attractive for end-device deployment because only a small subset of experts is active per token, bu

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

EditCLEVR: A Paired-Scene Intervention Benchmark for Compositional Faithfulness of Object-Centric Representations

DGX agent

arXiv:2607.22705v1 Announce Type: new Abstract: Object-centric learning aims to represent scenes as objects whose properties can be reused in new combinations. Existing evaluations usually score segme

model-releasesarxiv-cs-cv
28 Jul 2026
Research

Explaining BiomedCLIP with Weighted Banzhaf Interactions Supported by Tree-Gram Parsing

DGX agent

arXiv:2607.23368v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) are demonstrating significant capabilities in medical tasks like radiology analysis, yet providing faithful and interpre

researcharxiv-cs-ai
28 Jul 2026
Model Releases

Fairness Interventions in Classification: A Study on AI Explainability

DGX agent

arXiv:2407.14766v4 Announce Type: replace-cross Abstract: This paper presents a philosophical and experimental study of fairness interventions in AI classification, centered on the explainability and

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Flash-CNNCap: Capacitance Extraction via Image Mapping

DGX agent

arXiv:2607.23877v1 Announce Type: new Abstract: We present Flash-CNNCap, a CNN-based capacitance extractor that reformulates full-matrix capacitance prediction as image-to-image regression over spatia

model-releasesarxiv-cs-lg
28 Jul 2026
← Previous
1…342343344345346…1065
Next →