AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

Content type
87,042Total entries
1Added by human
87,041Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,106 results
Model Releases

TRU: Targeted Reverse Update for Efficient Multimodal Recommendation Unlearning

DGX agent

arXiv:2604.02183v2 Announce Type: replace Abstract: Multimodal recommendation systems (MRS) jointly model user-item interaction graphs and rich item content, but this tight coupling makes user data di

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

ADAG: Automatically Describing Attribution Graphs

DGX agent

arXiv:2604.07615v1 Announce Type: new Abstract: In language model interpretability research, extbf{circuit tracing} aims to identify which internal features causally contributed to a particular outp

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

AgentOpt v0.1 Technical Report: Client-Side Optimization for LLM-Based Agent

DGX agent

arXiv:2604.06296v1 Announce Type: cross Abstract: AI agents are increasingly deployed in real-world applications, including systems such as Manus, OpenClaw, and coding agents. Existing research has pr

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

AVGen-Bench: A Task-Driven Benchmark for Multi-Granular Evaluation of Text-to-Audio-Video Generation

DGX agent

arXiv:2604.08540v1 Announce Type: cross Abstract: Text-to-Audio-Video (T2AV) generation is rapidly becoming a core interface for media creation, yet its evaluation remains fragmented. Existing benchma

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Chunks as Arms: Multi-Armed Bandit-Guided Sampling for Long-Context LLM Preference Optimization

DGX agent

arXiv:2508.13993v2 Announce Type: replace Abstract: Long-context modeling is critical for a wide range of real-world tasks, including long-context question answering, summarization, and complex reason

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

ClawsBench: Evaluating Capability and Safety of LLM Productivity Agents in Simulated Workspaces

DGX agent

arXiv:2604.05172v2 Announce Type: replace Abstract: Large language model (LLM) agents are increasingly deployed to automate productivity tasks (e.g., email, scheduling, document management), but evalu

model-releasesarxiv-cs-ai
10 Apr 2026
Tutorials

Cross-Modal Emotion Transfer for Emotion Editing in Talking Face Video

DGX agent

arXiv:2604.07786v1 Announce Type: new Abstract: Talking face generation has gained significant attention as a core application of generative models. To enhance the expressiveness and realism of synthe

tutorialsarxiv-cs-cv
10 Apr 2026
Safety

Drift-Based Policy Optimization: Native One-Step Policy Learning for Online Robot Control

DGX agent

arXiv:2604.03540v2 Announce Type: replace Abstract: Although multi-step generative policies achieve strong performance in robotic manipulation by modeling multimodal action distributions, they require

safetyarxiv-cs-ro
10 Apr 2026
Model Releases

DSCA: Dynamic Subspace Concept Alignment for Lifelong VLM Editing

DGX agent

arXiv:2604.07965v1 Announce Type: new Abstract: Model editing aims to update knowledge to add new concepts and change relevant information without retraining. Lifelong editing is a challenging task, p

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

EditCaption: Human-Aligned Instruction Synthesis for Image Editing via Supervised Fine-Tuning and Direct Preference Optimization

DGX agent

arXiv:2604.08213v1 Announce Type: new Abstract: High-quality training triplets (source-target image pairs with precise editing instructions) are a critical bottleneck for scaling instruction-guided im

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Efficient and Effective Internal Memory Retrieval for LLM-Based Healthcare Prediction

DGX agent

arXiv:2604.07659v1 Announce Type: new Abstract: Large language models (LLMs) hold significant promise for healthcare, yet their reliability in high-stakes clinical settings is often compromised by hal

model-releasesarxiv-cs-cl
10 Apr 2026
Research

Efficient PRM Training Data Synthesis via Formal Verification

DGX agent

arXiv:2505.15960v3 Announce Type: replace Abstract: Process Reward Models (PRMs) have emerged as a promising approach for improving LLM reasoning capabilities by providing process supervision over rea

researcharxiv-cs-cl
10 Apr 2026
Safety

Explainable AI to Improve Machine Learning Reliability for Industrial Cyber-Physical Systems

DGX agent

arXiv:2601.16074v2 Announce Type: replace Abstract: Industrial Cyber-Physical Systems (CPS) are sensitive infrastructure from both safety and economics perspectives, making their reliability criticall

safetyarxiv-cs-lg
10 Apr 2026
Model Releases

FIT: A Large-Scale Dataset for Fit-Aware Virtual Try-On

DGX agent

arXiv:2604.08526v1 Announce Type: new Abstract: Given a person and a garment image, virtual try-on (VTO) aims to synthesize a realistic image of the person wearing the garment, while preserving their

model-releasesarxiv-cs-cv
10 Apr 2026
Research

Flemme: A Flexible and Modular Learning Platform for Medical Images

DGX agent

arXiv:2408.09369v3 Announce Type: replace-cross Abstract: As the rapid development of computer vision and the emergence of powerful network backbones and architectures, the application of deep learnin

researcharxiv-cs-cv
10 Apr 2026
Applications

GenLCA: 3D Diffusion for Full-Body Avatars from In-the-Wild Videos

DGX agent

arXiv:2604.07273v2 Announce Type: replace Abstract: We present GenLCA, a diffusion-based generative model for generating and editing photorealistic full-body avatars from text and image inputs. The ge

applicationsarxiv-cs-cv
10 Apr 2026
Applications

HEX: Humanoid-Aligned Experts for Cross-Embodiment Whole-Body Manipulation

DGX agent

arXiv:2604.07993v1 Announce Type: new Abstract: Humans achieve complex manipulation through coordinated whole-body control, whereas most Vision-Language-Action (VLA) models treat robot body parts larg

applicationsarxiv-cs-ro
10 Apr 2026
Model Releases

Holistic Optimal Label Selection for Robust Prompt Learning under Partial Labels

DGX agent

arXiv:2604.06614v1 Announce Type: cross Abstract: Prompt learning has gained significant attention as a parameter-efficient approach for adapting large pre-trained vision-language models to downstream

model-releasesarxiv-cs-lg
10 Apr 2026
Research

Lang2Act: Fine-Grained Visual Reasoning through Self-Emergent Linguistic Toolchains

DGX agent

arXiv:2602.13235v2 Announce Type: replace-cross Abstract: Visual Retrieval-Augmented Generation (VRAG) enhances Vision-Language Models (VLMs) by incorporating external visual documents to address a gi

researcharxiv-cs-cv
10 Apr 2026
Safety

Limits of Difficulty Scaling: Hard Samples Yield Diminishing Returns in GRPO-Tuned SLMs

DGX agent

arXiv:2604.06298v1 Announce Type: new Abstract: Recent alignment work on Large Language Models (LLMs) suggests preference optimization can improve reasoning by shifting probability mass toward better

safetyarxiv-cs-lg
10 Apr 2026
Research

LongSpec: Long-Context Lossless Speculative Decoding with Efficient Drafting and Verification

DGX agent

arXiv:2502.17421v4 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) can now process extremely long contexts, efficient inference over these extended inputs has become increasingl

researcharxiv-cs-ai
10 Apr 2026
Model Releases

LoRA-DA: Data-Aware Initialization for Low-Rank Adaptation via Asymptotic Analysis

DGX agent

arXiv:2510.24561v2 Announce Type: replace-cross Abstract: LoRA has become a widely adopted method for PEFT, and its initialization methods have attracted increasing attention. However, existing method

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

MedConclusion: A Benchmark for Biomedical Conclusion Generation from Structured Abstracts

DGX agent

arXiv:2604.06505v1 Announce Type: cross Abstract: Large language models (LLMs) are widely explored for reasoning-intensive research tasks, yet resources for testing whether they can infer scientific c

model-releasesarxiv-cs-ai
10 Apr 2026
Research

Novel View Synthesis as Video Completion

DGX agent

arXiv:2604.08500v1 Announce Type: new Abstract: We tackle the problem of sparse novel view synthesis (NVS) using video diffusion models; given K (approx 5) multi-view images of a scene and their

researcharxiv-cs-cv
10 Apr 2026
Model Releases

Open-Ended Instruction Realization with LLM-Enabled Multi-Planner Scheduling in Autonomous Vehicles

DGX agent

arXiv:2604.08031v1 Announce Type: cross Abstract: Most Human-Machine Interaction (HMI) research overlooks the maneuvering needs of passengers in autonomous driving (AD). Natural language offers an int

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

PASK: Toward Intent-Aware Proactive Agents with Long-Term Memory

DGX agent

arXiv:2604.08000v1 Announce Type: cross Abstract: Proactivity is a core expectation for AGI. Prior work remains largely confined to laboratory settings, leaving a clear gap in real-world proactive age

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

PhyEdit: Towards Real-World Object Manipulation via Physically-Grounded Image Editing

DGX agent

arXiv:2604.07230v2 Announce Type: replace Abstract: Achieving physically accurate object manipulation in image editing is essential for its potential applications in interactive world models. However,

model-releasesarxiv-cs-cv
10 Apr 2026
Research

Reasoning Fails Where Step Flow Breaks

DGX agent

arXiv:2604.06695v1 Announce Type: new Abstract: Large reasoning models (LRMs) that generate long chains of thought now perform well on multi-step math, science, and coding tasks. However, their behavi

researcharxiv-cs-ai
10 Apr 2026
Model Releases

REVEAL: Reasoning-Enhanced Forensic Evidence Analysis for Explainable AI-Generated Image Detection

DGX agent

arXiv:2511.23158v2 Announce Type: replace-cross Abstract: The rapid progress of visual generative models has made AI-generated images increasingly difficult to distinguish from authentic ones, posing

model-releasesarxiv-cs-ai
10 Apr 2026
Safety

RewardFlow: Generate Images by Optimizing What You Reward

DGX agent

arXiv:2604.08536v1 Announce Type: new Abstract: We introduce RewardFlow, an inversion-free framework that steers pretrained diffusion and flow-matching models at inference time through multi-reward La

safetyarxiv-cs-cv
10 Apr 2026
Model Releases

SearchAD: Large-Scale Rare Image Retrieval Dataset for Autonomous Driving

DGX agent

arXiv:2604.08008v1 Announce Type: new Abstract: Retrieving rare and safety-critical driving scenarios from large-scale datasets is essential for building robust autonomous driving (AD) systems. As dat

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

SpecQuant: Spectral Decomposition and Adaptive Truncation for Ultra-Low-Bit LLMs Quantization

DGX agent

arXiv:2511.11663v2 Announce Type: replace-cross Abstract: The emergence of accurate open large language models (LLMs) has sparked a push for advanced quantization techniques to enable efficient deploy

model-releasesarxiv-cs-ai
10 Apr 2026
Research

SVGFusion: A VAE-Diffusion Transformer for Vector Graphic Generation

DGX agent

arXiv:2412.10437v3 Announce Type: replace Abstract: Generating high-quality Scalable Vector Graphics (SVGs) from text remains a significant challenge. Existing LLM-based models that generate SVG code

researcharxiv-cs-cv
10 Apr 2026
Model Releases

Visually-grounded Humanoid Agents

DGX agent

arXiv:2604.08509v1 Announce Type: new Abstract: Digital human generation has been studied for decades and supports a wide range of real-world applications. However, most existing systems are passively

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

When Personalization Tricks Detectors: The Feature-Inversion Trap in Machine-Generated Text Detection

DGX agent

arXiv:2510.12476v2 Announce Type: replace Abstract: Large language models (LLMs) have grown more powerful in language generation, producing fluent text and even imitating personal style. Yet, this abi

model-releasesarxiv-cs-cl
10 Apr 2026
Safety

Beyond the Trace: Coupling an Interpretable Reasoning-State Readout to Native MoE Routing

DGX agent

arXiv:2608.17638v1 Announce Type: new Abstract: What a reasoning model writes is only a partial record of the process that produces it. We introduce a two-level internal readout for mixture-of-experts

safetyarxiv-cs-ai
19 Aug 2026
Model Releases

Improving Complex Moire Removal with Generative Supervision

DGX agent

arXiv:2608.17883v1 Announce Type: new Abstract: The availability of high-quality paired data is essential for training learning-based image demoireing models. However, it remains challenging for exist

model-releasesarxiv-cs-cv
19 Aug 2026
Model Releases

Key-Frame Reasoning with SAM3: Third Place Solution for the MeViS-Text Track of the 8th LSVOS Challenge

DGX agent

arXiv:2608.17279v1 Announce Type: new Abstract: This report presents a two-stage, training-free solution for the MeViS-Text track of the 8th LSVOS Challenge. The task requires a model to localize and

model-releasesarxiv-cs-cv
19 Aug 2026
Model Releases

Leveraging existing sparse point annotations for benthic imagery dense segmentation

DGX agent

arXiv:2608.17561v1 Announce Type: new Abstract: The health of marine ecosystems is a critical indicator of global environmental change, yet the physical constraints of underwater observation and the i

model-releasesarxiv-cs-cv
19 Aug 2026
Model Releases

MANIGUARD: A Benchmark and Data Suite for Specification-Grounded Safety Evaluation and Improvement of Robotic Manipulation

DGX agent

arXiv:2608.17386v1 Announce Type: new Abstract: Foundation-model policies for robotic manipulation are advancing rapidly on task success, but rigorous evaluation of whether they succeed safely is stil

model-releasesarxiv-cs-ro
19 Aug 2026
Local Ai

Mixture-of-Expert Blocks Contain Strong Hallucination Detection Signals

DGX agent

arXiv:2608.17687v1 Announce Type: new Abstract: Despite their widespread use, Large Language Models (LLMs) remain limited by a fundamental problem: the generation of plausible but false content, known

local-aiarxiv-cs-ai
19 Aug 2026
Safety

Policy-Invariant Reward Shaping from LLM Feedback: A Framework for Hybrid RL Agents

DGX agent

arXiv:2608.18008v1 Announce Type: cross Abstract: Combining large language models with reinforcement learning is increasingly explored, yet the theoretical status of LLM-derived reward signals is ofte

safetyarxiv-cs-ai
19 Aug 2026
Model Releases

Probing the Prefill: Detecting Code Vulnerabilities via Latent Activations

DGX agent

arXiv:2608.16970v1 Announce Type: cross Abstract: LLM-based code generation is now embedded in mission-critical pipelines, but defenses against vulnerable output remain post-hoc -- static analyzers, f

model-releasesarxiv-cs-ai
19 Aug 2026
Model Releases

PTXBench: Benchmark and Adapt LLMs for GPU Kernel Optimization with Architecture-specific PTX

DGX agent

arXiv:2608.17379v1 Announce Type: cross Abstract: We introduce PTXBench, a benchmark for evaluating and adapting large language models (LLMs) to use architecture-specific PTX for GPU kernel optimizati

model-releasesarxiv-cs-ai
19 Aug 2026
Model Releases

Q-Interference: Memory-Efficient Phase-Aware Quantum-Inspired Attention

DGX agent

arXiv:2608.17288v1 Announce Type: new Abstract: GPT attention measures token compatibility through dot-product similarity. This mechanism is simple, effective, and memory-efficient. But it does not ex

model-releasesarxiv-cs-cl
19 Aug 2026
Model Releases

S^3AM: A Single-Stream SAM with Reliability-Calibrated Frequency Adapter for Multi-modal Salient Object Detection

DGX agent

arXiv:2608.17475v1 Announce Type: new Abstract: Vision foundation models have recently advanced multi-modal salient object detection (MSOD) through parameter-efficient tuning and prompt learning. Howe

model-releasesarxiv-cs-cv
19 Aug 2026
Safety

Teach and Grow: An Agent-Centered Architecture for General Robot Learning

DGX agent

arXiv:2608.17209v1 Announce Type: cross Abstract: End-to-end vision-language-action (VLA) and world-action models offer an elegant route to general-purpose robotics, but their reliability is bounded b

safetyarxiv-cs-ai
19 Aug 2026
Research

Towards Safer RAG: Only Agents Capable of System 2 Thinking may Access Untrusted Documents

DGX agent

arXiv:2608.17153v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has significantly enhanced the performance of large language models (LLMs), yet these systems remain vulnerable to

researcharxiv-cs-cl
19 Aug 2026
← Previous
1…334335336337338…1065
Next →