AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,965
  • Agents7,446
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,131
  • Local Ai4,857
  • Model Releases23,360
  • Research19,834
  • Safety13,174
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,965
  • Agents7,446
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,131
  • Local Ai4,857
  • Model Releases23,360
  • Research19,834
  • Safety13,174
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlog
86,965Total entries
1Added by human
86,964Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,458 results
Safety

EPiC: Efficient Video Camera Control Learning with Precise Anchor-Video Guidance

DGX agent

arXiv:2505.21876v2 Announce Type: replace-cross Abstract: Recent approaches for video generation with camera control often create anchor videos (i.e., rendered videos that approximate desired camera m

safetyarxiv-cs-ai
29 May 2026
Research
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

EvA: An Evidence-First Audio Understanding Paradigm for LALMs

DGX agent

arXiv:2603.27667v2 Announce Type: replace-cross Abstract: Large Audio Language Models (LALMs) still struggle in complex acoustic scenes because they often fail to preserve task-relevant acoustic evide

researcharxiv-cs-ai
29 May 2026
Applications

Evaluation of Conversational Agents: Understanding Culture, Context and Environment in Emotion Detection

DGX agent

arXiv:2605.30099v1 Announce Type: new Abstract: Valuable decisions and highly prioritized analysis now depend on applications such as facial biometrics, social media photo tagging, and human robots in

applicationsarxiv-cs-cv
29 May 2026
Safety

Evolutionary Refinement of Generative Graph Topologies: A Hybrid WGAN-GA Approach

DGX agent

arXiv:2605.29161v1 Announce Type: cross Abstract: Generating realistic graph-structured data is challenging due to discrete connectivity, varying graph sizes, and class-specific structural patterns. R

safetyarxiv-cs-ai
29 May 2026
Tutorials

Feedback-to-Rubrics: Can We Learn Expert Criteria from Inline Comments?

DGX agent

arXiv:2605.29857v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for writing and review support, but their usefulness depends on context-dependent criteria, such as e

tutorialsarxiv-cs-lg
29 May 2026
Safety

Fisher-Preserving Guidance: Training-Free Manifold Constraints for Safe Diffusion Control

DGX agent

arXiv:2605.29937v1 Announce Type: cross Abstract: Diffusion models are effective for waypoint prediction in visual navigation, but standard sampling and test time guidance can produce unreliable or in

safetyarxiv-cs-lg
29 May 2026
Safety

FlowSeg: Dynamic Semantic Guidance for LLM-Conditioned Segmentation

DGX agent

arXiv:2605.29461v1 Announce Type: new Abstract: LLM-conditioned segmentation has recently advanced rapidly by coupling large language models with iterative mask generation frameworks. However, we iden

safetyarxiv-cs-cv
29 May 2026
Tools

Function invocations now billed per unit

DGX agent

Vercel updated its billing model to charge for serverless function invocations on a per-unit basis rather than previous pricing structures. This change affects how customers are charged for executing

toolsvercel-blog
29 May 2026
Safety

GAPD: Gold-Action Policy Distillation for Agentic Reinforcement Learning in Knowledge Base Question Answering

DGX agent

arXiv:2605.29584v1 Announce Type: new Abstract: Reinforcement learning (RL) is a natural fit for agentic knowledge base question answering (KBQA), where a model must issue executable actions, observe

safetyarxiv-cs-cl
29 May 2026
Safety

GASS: Geometry-Aware Spherical Sampling for Disentangled Diversity Enhancement in Text-to-Image Generation

DGX agent

arXiv:2602.17200v2 Announce Type: replace Abstract: Despite high semantic alignment, modern text-to-image (T2I) generative models still struggle to synthesize diverse images from a given prompt. In th

safetyarxiv-cs-cv
29 May 2026
Safety

Gaze2Act: Gaze-Conditioned Vision-Language-Action Policies for Interactive Robot Manipulation

DGX agent

arXiv:2605.30282v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have recently shown strong potential for robot learning by following language instructions. However, in practice, la

safetyarxiv-cs-ro
29 May 2026
Applications

Generative Spatiotemporal Intent Sequence Recommendation via Implicit Reasoning in Amap

DGX agent

arXiv:2605.28888v1 Announce Type: cross Abstract: Real-world user behavior rarely consists of isolated actions; instead, it often forms intent flows governed by spatiotemporal dependencies. To provide

applicationsarxiv-cs-lg
29 May 2026
Safety

Genetically Aligned Patient Representations Improve Hematological Diagnosis

DGX agent

arXiv:2605.29980v1 Announce Type: cross Abstract: Multimodal alignment of histopathology encoders with transcriptomic and genomic data has been shown to significantly improve performance in downstream

safetyarxiv-cs-ai
29 May 2026
Safety

GrepSeek: Training Search Agents for Direct Corpus Interaction

DGX agent

arXiv:2605.29307v1 Announce Type: cross Abstract: Large Language Model (LLM) search agents have shown strong promise for knowledge-intensive language tasks through multiple rounds of reasoning and inf

safetyarxiv-cs-ai
29 May 2026
Agents

Hijacking Agent Memory: Stealthy Trojan Attacks Through Conversational Interaction

DGX agent

arXiv:2605.29960v1 Announce Type: cross Abstract: Large language model (LLM) agents increasingly leverage long term memory to support persistent and autonomous task execution. However, this capability

agentsarxiv-cs-ai
29 May 2026
Agents

How Consistent Are LLM Agents? Measuring Behavioral Reproducibility in Multi-Step Tool-Calling Pipelines

DGX agent

arXiv:2605.28840v1 Announce Type: cross Abstract: Large language model (LLM) agents with tool-calling capabilities are increasingly deployed in production systems, yet a fundamental reliability questi

agentsarxiv-cs-ai
29 May 2026
Hardware

How Together AI built the world’s fastest speech-to-text stack

DGX agent

Together AI developed an optimized speech-to-text system focused on achieving the fastest processing speeds through technical innovations in their inference stack and model optimization. The approach

hardwaretogether-ai-blog
29 May 2026
Safety

Inferring Code Correctness from Specification

DGX agent

arXiv:2605.29822v1 Announce Type: cross Abstract: Large language models (LLMs) have become integral to modern software development, enabling automated code generation at scale. However, validating the

safetyarxiv-cs-ai
29 May 2026
Research

IP-Adapter Is All You Need: Towards Fine-Tuning-Free Diffusion-Based Talking Face Generation

DGX agent

arXiv:2605.30230v1 Announce Type: new Abstract: With the rapid advancement of diffusion models, talking face generation has made remarkable progress. However, existing diffusion-based methods still re

researcharxiv-cs-cv
29 May 2026
Local Ai

KAN-AD: Time Series Anomaly Detection with Kolmogorov-Arnold Networks

DGX agent

arXiv:2411.00278v4 Announce Type: replace Abstract: Time series anomaly detection (TSAD) underpins real-time monitoring in cloud services and web systems, allowing rapid identification of anomalies to

local-aiarxiv-cs-lg
29 May 2026
Tutorials

Knowing What to Solve Before How: Preplan Empowered LLM Mathematical Reasoning

DGX agent

arXiv:2605.30245v1 Announce Type: new Abstract: Current plan-based reasoning methods improve large language models (LLMs) by inserting a planning stage before execution, giving rise to the question ri

tutorialsarxiv-cs-cl
29 May 2026
Tutorials

Learning and Adaptation in Wire Arc Additive Manufacturing Bead Geometry Control

DGX agent

arXiv:2605.29144v1 Announce Type: new Abstract: Robotics Wire Arc Additive Manufacturing (WAAM) is governed by complex and nonlinear process dynamics coupling thermal field to the build geometry. The

tutorialsarxiv-cs-ro
29 May 2026
Research

Learning Context-Conditioned Predicate Semantics via Prototype Feedback

DGX agent

arXiv:2605.29610v1 Announce Type: cross Abstract: In scene graph generation, a central challenge is modeling polysemous predicates whose meanings shift across contexts. Prior approaches address this i

researcharxiv-cs-ai
29 May 2026
Tutorials

Learning Robust and Task-Invariant Functional Representation from fMRI through Siamese Self-Supervised Learning

DGX agent

arXiv:2605.28990v1 Announce Type: new Abstract: Functional magnetic resonance imaging (fMRI) is a powerful tool for investigating human brain function. However, the high cost of data acquisition and t

tutorialsarxiv-cs-lg
29 May 2026
Research

Lightweight Complementary-Cue Fusion for Robust Video Face Forgery Detection

DGX agent

arXiv:2605.29092v1 Announce Type: new Abstract: Current face video forgery detectors use wide or dual-stream backbones. We show that a single, lightweight fusion of two handcrafted cues can achieve hi

researcharxiv-cs-cv
29 May 2026
Research

Matching Rates and Optimal Allocation for Federated Probe-Logit Distillation under Heterogeneous Bandwidth Budgets

DGX agent

arXiv:2605.29642v1 Announce Type: cross Abstract: In federated language modeling, K nodes each hold n samples but cannot pool data or exchange full-precision gradients or weights. We study the minimax

researcharxiv-cs-lg
29 May 2026
Research

MedCase-Structured: A Text-to-FHIR Dataset for Benchmarking Diagnostic Reasoning in Clinically Realistic EHR Settings

DGX agent

arXiv:2605.30295v1 Announce Type: cross Abstract: Large language models (LLMs) show promise for clinical reasoning and decision support, but evaluation in realistic, electronic health record-congruent

researcharxiv-cs-ai
29 May 2026
Local Ai

MediHive: A Decentralized Agent Collective for Medical Reasoning

DGX agent

arXiv:2603.27150v2 Announce Type: replace Abstract: Large language models (LLMs) have revolutionized medical reasoning tasks, yet single-agent systems often falter on complex, interdisciplinary proble

local-aiarxiv-cs-ai
29 May 2026
Safety

MetaRanker: Human-in-the-loop Active Ranking for Metalens Image Quality

DGX agent

arXiv:2605.29212v1 Announce Type: new Abstract: Image quality in modern imaging systems emerges from the coupled effects of the sensor, optics, and computational reconstruction. Ultra-thin metalenses

safetyarxiv-cs-cv
29 May 2026
Safety

Metric-Dependent Annotation Saturation for Learning from Label Distributions

DGX agent

arXiv:2605.29797v1 Announce Type: new Abstract: When annotators disagree on a label, the disagreement itself carries signal -- and the number of annotators needed to capture it depends on the evaluati

safetyarxiv-cs-cl
29 May 2026
Safety

Mining or Synthesis? Rethinking Exploration Efficiency in Iterative Alignment of Mathematical Reasoning

DGX agent

arXiv:2602.05370v3 Announce Type: replace Abstract: Iterative Direct Preference Optimization (DPO) has emerged as a widely used paradigm for aligning Large Language Models on reasoning tasks. Existing

safetyarxiv-cs-cl
29 May 2026
Applications

MOO: A Multi-view Oriented Observations Dataset for Viewpoint Analysis in Cattle Re-Identification

DGX agent

arXiv:2603.04314v2 Announce Type: replace-cross Abstract: Animal re-identification (ReID) faces critical challenges due to viewpoint variations, particularly in Aerial-Ground (AG-ReID) settings where

applicationsarxiv-cs-ai
29 May 2026
Agents

MOOSE-Copilot: A Web-Based Interactive Assistant for Unified Exploratory and Fine-Grained Scientific Hypothesis Discovery

DGX agent

arXiv:2605.29475v1 Announce Type: cross Abstract: Large language models (LLMs) show remarkable potential in scientific hypothesis discovery. However, existing approaches face two critical limitations:

agentsarxiv-cs-ai
29 May 2026
Research

Neural-Behavioral Representation of Natural Whole-body Movement in Monkeys

DGX agent

arXiv:2605.29355v1 Announce Type: new Abstract: Understanding how cortical activity represents natural whole-body behaviors in primates remains challenging. Limited by the diversity of movements and i

researcharxiv-cs-lg
29 May 2026
Research

OccamToken: Efficient VLM Inference with Training-Free and Budget-Adaptive Token Pruning

DGX agent

arXiv:2605.29657v1 Announce Type: cross Abstract: Vision-language models (VLMs) rely on long visual token sequences for visual understanding, making the prefill stage expensive in both computation and

researcharxiv-cs-ai
29 May 2026
Tutorials

OmniAID: Decoupling Semantic and Artifacts for Universal AI-Generated Image Detection in the Wild

DGX agent

arXiv:2511.08423v3 Announce Type: replace Abstract: A truly universal AI-Generated Image (AIGI) detector must simultaneously generalize across diverse generative models and varied semantic content. Cu

tutorialsarxiv-cs-cv
29 May 2026
Tutorials

One Mask to Rule Them All: On Hidden Facts after Editing and How to Find Them

DGX agent

arXiv:2605.28839v1 Announce Type: new Abstract: Knowledge editing methods such as ROME and MEMIT update factual associations in transformer models by modifying MLP weights. While evaluated mainly by o

tutorialsarxiv-cs-lg
29 May 2026
Safety

Orthogonal Negative Guidance in Attention Feature Space for Text-to-Image Generation

DGX agent

arXiv:2605.29390v1 Announce Type: new Abstract: Text-to-image (T2I) models have become increasingly capable of generating high-quality images. Yet, enforcing the explicit absence of a specified object

safetyarxiv-cs-cv
29 May 2026
Agents

Parse PDFs at lightspeed (this video is at 1x) Absolute cinema

DGX agent

Parse PDFs at lightspeed (this video is at 1x) Absolute cinema Media We've created the world's fastest PDF parser ⚡️ And it's more accurate than any other open-source, model-free PDF parser out there

agentsjerry-liu--x
29 May 2026
Safety

PersonaAgent: Bridging Memory and Action for Personalized LLM Agents

DGX agent

arXiv:2506.06254v2 Announce Type: replace Abstract: Large Language Model (LLM) empowered agents have recently emerged as advanced paradigms that exhibit impressive capabilities in a wide range of doma

safetyarxiv-cs-ai
29 May 2026
Safety

RL2ML: Finite-Rollout Surrogate Objectives from Reinforcement Learning to Maximum Likelihood

DGX agent

arXiv:2605.30154v1 Announce Type: new Abstract: Correctness-based Reinforcement Learning with Verifiable Rewards (RLVR) trains language models from binary feedback on sampled outputs, but the objectiv

safetyarxiv-cs-lg
29 May 2026
Safety

Rooted Absorbed Prefix Trajectory Balance with Submodular Replay for GFlowNet Training

DGX agent

arXiv:2603.00454v2 Announce Type: replace-cross Abstract: Generative Flow Networks (GFlowNets) enable fine-tuning large language models to approximate reward-proportional posteriors, but they remain p

safetyarxiv-cs-ai
29 May 2026
Hardware

Run Step 3.7 Flash on NVIDIA GPUs with Enterprise-Ready Multimodal AI

DGX agent

Step 3.7 Flash is a 198B-parameter Mixture-of-Experts vision-language model designed for enterprise-scale production workloads, featuring native image and video input, a 256k context window, and confi

hardwarenvidia-developer
29 May 2026
Research

S2MDF: A Plug-And-Play Layer for Intersection-Free Multi-Object Signed Distance Fields

DGX agent

arXiv:2605.29761v1 Announce Type: new Abstract: Compositional implicit surface representations model scenes as collections of objects, each encoded by a Signed Distance Field (SDF). A fundamental limi

researcharxiv-cs-cv
29 May 2026
Safety

SafeRx-Agent: A Knowledge-Grounded Multi-Agent Framework for Safe and Explainable Medication Recommendation

DGX agent

arXiv:2605.29146v1 Announce Type: cross Abstract: Medication recommendation predicts medications for patient visits, but existing methods still face two key challenges. At the model level, traditional

safetyarxiv-cs-ai
29 May 2026
Research

SAVAA: Mitigating Hallucinations in LVLMs via Step-wise Adaptive Visual Attention Amplification

DGX agent

arXiv:2602.13600v2 Announce Type: replace Abstract: A line of recent training-free methods for mitigating hallucinations in large vision-language models (LVLMs) operates by amplifying attention to vis

researcharxiv-cs-cv
29 May 2026
Agents

SEAL: Can Saturated Benchmarks Be Revived by LLM-as-a-Meta-Judge?

DGX agent

arXiv:2605.30104v1 Announce Type: new Abstract: Widely used language-model benchmarks are increasingly saturated, with frontier systems often receiving near-tied scores that standard metrics cannot re

agentsarxiv-cs-cl
29 May 2026
Research

Solving Integer Linear Programming with Parallel Tempering

DGX agent

arXiv:2605.29366v1 Announce Type: new Abstract: Integer Linear Programming (ILP) serves as a versatile framework for modeling a wide range of combinatorial optimization problems, typically addressed b

researcharxiv-cs-lg
29 May 2026
← Previous
1…974975976977978…1302
Next →