AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlog
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,794 results
Research

FrameDiT: Diffusion Transformer with Matrix Attention for Efficient Video Generation

DGX agent

arXiv:2603.09721v2 Announce Type: replace Abstract: High-fidelity video generation remains challenging for diffusion models due to the difficulty of modeling complex spatio-temporal dynamics efficient

researcharxiv-cs-cv
21 Apr 2026
Research
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Frequency-guided Multi-level Reasoning for Scene Graph Generation in Video

DGX agent

arXiv:2604.17298v1 Announce Type: new Abstract: Video Scene Graph Generation aims to obtain structured semantic representations of objects and their relationships in videos for high-level understandin

researcharxiv-cs-cv
21 Apr 2026
Research

From 2:4 to 8:16 sparsity patterns in LLMs for Outliers and Weights with Variance Correction

DGX agent

arXiv:2507.03052v2 Announce Type: replace Abstract: As large language models (LLMs) grow in size, efficient compression techniques like quantization and sparsification are critical. While quantization

researcharxiv-cs-lg
21 Apr 2026
Model Releases

GeoRC: A Benchmark for Geolocation Reasoning Chains

DGX agent

arXiv:2601.21278v2 Announce Type: replace-cross Abstract: Vision Language Models (VLMs) are good at recognizing the global location of a photograph -- their geolocation prediction accuracy rivals the

model-releasesarxiv-cs-cl
21 Apr 2026
Agents

Harness Engineering Without the Hype: 🦄 AI That Works #54 https://x.com/i/broadcasts/1DxLdvQrDwkxm

DGX agent

This episode of 'AI That Works' discusses practical approaches to harness engineering in AI systems, likely focusing on techniques for effectively prompting and controlling AI model behavior beyond ma

agentsharrison-chase--x
21 Apr 2026
Safety

High-Resolution Visual Reasoning via Multi-Turn Grounding-Based Reinforcement Learning

DGX agent

arXiv:2507.05920v2 Announce Type: replace Abstract: State-of-the-art large multi-modal models (LMMs) face challenges when processing high-resolution images, as these inputs are converted into enormous

safetyarxiv-cs-cv
21 Apr 2026
Research

HopWeaver: Cross-Document Synthesis of High-Quality and Authentic Multi-Hop Questions

DGX agent

arXiv:2505.15087v3 Announce Type: replace Abstract: Multi-Hop Question Answering (MHQA) is crucial for evaluating the model's capability to integrate information from diverse sources. However, creatin

researcharxiv-cs-cl
21 Apr 2026
Agents

How to Ground a Korean AI Agent in Real Demographics with Synthetic Personas

DGX agent

This article explains how to ground Korean AI agents in realistic demographic data by utilizing synthetic personas, likely leveraging NVIDIA's Nemotron models available on Hugging Face. The approach e

agentshugging-face
21 Apr 2026
Model Releases

HUGGING FACE JUST AUTOMATED THEIR ENTIRE POST-TRAINING TEAM WITH AN AGENT. It reads papers, runs GPU experiments, iterates, and builds resea…

DGX agent

HUGGING FACE JUST AUTOMATED THEIR ENTIRE POST-TRAINING TEAM WITH AN AGENT. It reads papers, runs GPU experiments, iterates, and builds research-backed models autonomously. Pushed a benchmark from 10%

model-releasesclem-delangue--x
21 Apr 2026
Research

IMA-MoE: An Interpretable Modality-Aware Mixture-of-Experts Framework for Characterizing the Neurobiological Signatures of Binge Eating Disorder

DGX agent

arXiv:2604.17028v1 Announce Type: new Abstract: Binge eating disorder (BED) is the most prevalent eating disorder. However, current diagnostic frameworks remain largely grounded in symptom-based crite

researcharxiv-cs-cv
21 Apr 2026
Safety

Implicit neural representations as a coordinate-based framework for continuous environmental field reconstruction from sparse ecological observations

DGX agent

arXiv:2604.18083v1 Announce Type: new Abstract: Reconstructing continuous environmental fields from sparse and irregular observations remains a central challenge in environmental modelling and biodive

safetyarxiv-cs-lg
21 Apr 2026
Model Releases

Inflated Excellence or True Performance? Rethinking Medical Diagnostic Benchmarks with Dynamic Evaluation

DGX agent

arXiv:2510.09275v2 Announce Type: replace Abstract: Medical diagnostics is a high-stakes and complex domain that is critical to patient care. However, current evaluations of large language models (LLM

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Kimi K2.6 is now live inside Anything!

DGX agent

Kimi K2.6, an AI model from Moonshot, has been integrated into the Anything platform. This update likely enables users to access Kimi's capabilities directly within the Anything application interface.

model-releaseskimi-moonshot--x
21 Apr 2026
Model Releases

LayerCache: Exploiting Layer-wise Velocity Heterogeneity for Efficient Flow Matching Inference

DGX agent

arXiv:2604.16492v1 Announce Type: new Abstract: Flow Matching models achieve state-of-the-art image generation quality but incur substantial inference cost due to iterative denoising through large Tra

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Learning Stable Predictors from Weak Supervision under Distribution Shift

DGX agent

arXiv:2604.05002v2 Announce Type: replace Abstract: Learning from weak, proxy, or relative supervision is common when ground-truth labels are unavailable, but robustness under distribution shift remai

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Learning to Control Summaries with Score Ranking

DGX agent

arXiv:2604.17197v1 Announce Type: new Abstract: Recent advances in summarization research focus on improving summary quality across multiple criteria, such as completeness, conciseness, and faithfulne

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Learning to Correct: Calibrated Reinforcement Learning for Multi-Attempt Chain-of-Thought

DGX agent

arXiv:2604.17912v1 Announce Type: new Abstract: State-of-the-art reasoning models utilize long chain-of-thought (CoT) to solve increasingly complex problems using more test-time computation. In this w

researcharxiv-cs-lg
21 Apr 2026
Model Releases

Learning to Retrieve User History and Generate User Profiles for Personalized Persuasiveness Prediction

DGX agent

arXiv:2601.05654v3 Announce Type: replace Abstract: Estimating the persuasiveness of messages is critical in various applications, from recommender systems to safety assessment of LLMs. While it is im

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Lil: Less is Less When Applying Post-Training Sparse-Attention Algorithms in Long-Decode Stage

DGX agent

arXiv:2601.03043v3 Announce Type: replace Abstract: Large language models (LLMs) demonstrate strong capabilities across a wide range of complex tasks and are increasingly deployed at scale, placing si

researcharxiv-cs-cl
21 Apr 2026
Model Releases

LIVE: Leveraging Image Manipulation Priors for Instruction-based Video Editing

DGX agent

arXiv:2604.17021v1 Announce Type: new Abstract: Video editing aims to modify input videos according to user intent. Recently, end-to-end training methods have garnered widespread attention, constructi

model-releasesarxiv-cs-cv
21 Apr 2026
Research

LLM Hypnosis: Exploiting User Feedback for Unauthorized Knowledge Injection to All Users

DGX agent

arXiv:2507.02850v3 Announce Type: replace Abstract: We describe a vulnerability in language models (LMs) trained with user feedback, whereby a single user can persistently alter LM knowledge and behav

researcharxiv-cs-cl
21 Apr 2026
Tutorials

Lorentz Framework for Semantic Segmentation

DGX agent

arXiv:2604.16836v1 Announce Type: new Abstract: Semantic segmentation in hyperbolic space enables compact modeling of hierarchical structure while providing inherent uncertainty quantification. Prior

tutorialsarxiv-cs-cv
21 Apr 2026
Model Releases

ltzGLUE: Luxembourgish General Language Understanding Evaluation

DGX agent

arXiv:2604.17976v1 Announce Type: new Abstract: This paper presents ltzGLUE, the first Natural Language Understanding (NLU) benchmark for Luxembourgish (LTZ) based on the popular GLUE benchmark for En

model-releasesarxiv-cs-cl
21 Apr 2026
Applications

MambaKick: Early Penalty Direction Prediction from HAR Embeddings

DGX agent

arXiv:2604.16588v1 Announce Type: new Abstract: Penalty kicks in soccer are decided under extreme time constraints, where goalkeepers benefit from anticipating shot direction from the kickers motion b

applicationsarxiv-cs-cv
21 Apr 2026
Model Releases

MemBuilder: Reinforcing LLMs for Long-Term Memory Construction via Attributed Dense Rewards

DGX agent

arXiv:2601.05488v3 Announce Type: replace Abstract: Maintaining consistency in long-term dialogues remains a fundamental challenge for LLMs, as standard retrieval mechanisms often fail to capture the

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Multimodal In-context Learning for ASR of Low-resource Languages

DGX agent

arXiv:2601.05707v2 Announce Type: replace Abstract: Automatic speech recognition (ASR) still covers only a small fraction of the world's languages, mainly due to supervised data scarcity. In-context l

model-releasesarxiv-cs-cl
21 Apr 2026
Research

MUSEG: Reinforcing Video Temporal Understanding via Timestamp-Aware Multi-Segment Grounding

DGX agent

arXiv:2505.20715v2 Announce Type: replace-cross Abstract: Video temporal understanding is crucial for multimodal large language models (MLLMs) to reason over events in videos. Despite recent advances

researcharxiv-cs-cl
21 Apr 2026
Local Ai

Network-wide Freeway Traffic Estimation Using Sparse Sensor Data: A Dirichlet Graph Auto-Encoder Approach

DGX agent

arXiv:2503.15845v2 Announce Type: replace Abstract: Network-wide Traffic State Estimation (TSE), which aims to infer a complete image of network traffic states with sparsely deployed sensors, plays a

local-aiarxiv-cs-lg
21 Apr 2026
Research

One-Step Diffusion with Inverse Residual Fields for Unsupervised Industrial Anomaly Detection

DGX agent

arXiv:2604.18393v1 Announce Type: new Abstract: Diffusion models have achieved outstanding performance in unsupervised industrial anomaly detection (uIAD) by learning a manifold of normal data under t

researcharxiv-cs-cv
21 Apr 2026
Model Releases

OPeRA: A Dataset of Observation, Persona, Rationale, and Action for Evaluating LLMs on Human Online Shopping Behavior Simulation

DGX agent

arXiv:2506.05606v5 Announce Type: replace Abstract: Can large language models (LLMs) accurately simulate the next web action of a specific user? While LLMs have shown promising capabilities in generat

model-releasesarxiv-cs-cl
21 Apr 2026
Local Ai

Privacy-R1: Privacy-Aware Multi-LLM Agent Collaboration via Reinforcement Learning

DGX agent

arXiv:2510.16054v2 Announce Type: replace-cross Abstract: When users submit queries to Large Language Models (LLMs), their prompts can often contain sensitive data, forcing a difficult choice: Send th

local-aiarxiv-cs-cl
21 Apr 2026
Model Releases

Progressive Online Video Understanding with Evidence-Aligned Timing and Transparent Decisions

DGX agent

arXiv:2604.18459v1 Announce Type: new Abstract: Visual agents operating in the wild must respond to queries precisely when sufficient evidence first appears in a video stream, a critical capability th

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Pseudo2Real: Task Arithmetic for Pseudo-Label Correction in Automatic Speech Recognition

DGX agent

arXiv:2510.08047v2 Announce Type: replace-cross Abstract: Robust ASR under domain shift is crucial because real-world systems encounter unseen accents and domains with limited labeled data. Although p

model-releasesarxiv-cs-cl
21 Apr 2026
Applications

ReconVLA: An Uncertainty-Guided and Failure-Aware Vision-Language-Action Framework for Robotic Control

DGX agent

arXiv:2604.16677v1 Announce Type: new Abstract: Vision-language-action (VLA) models have emerged as generalist robotic controllers capable of mapping visual observations and natural language instructi

applicationsarxiv-cs-ro
21 Apr 2026
Model Releases

ResearchBench: Benchmarking LLMs in Scientific Discovery via Inspiration-Based Task Decomposition

DGX agent

arXiv:2503.21248v3 Announce Type: replace Abstract: Large language models (LLMs) have shown potential in assisting scientific research, yet their ability to discover high-quality research hypotheses r

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Revisiting Active Sequential Prediction-Powered Mean Estimation

DGX agent

arXiv:2604.18569v1 Announce Type: cross Abstract: In this work, we revisit the problem of active sequential prediction-powered mean estimation, where at each round one must decide the query probabilit

model-releasesarxiv-cs-lg
21 Apr 2026
Research

Sampling for Quality: Training-Free Reward-Guided LLM Decoding via Sequential Monte Carlo

DGX agent

arXiv:2604.16453v1 Announce Type: new Abstract: We introduce a principled probabilistic framework for reward-guided decoding in large language models, addressing the limitations of standard decoding m

researcharxiv-cs-lg
21 Apr 2026
Model Releases

Scaling Codex to enterprises worldwide

DGX agent

OpenAI's Codex, a code generation model trained on publicly available code, was scaled for enterprise deployment to provide developers with AI-assisted coding capabilities across organizations worldwi

model-releasesopenai
21 Apr 2026
Model Releases

Scaling External Knowledge Input Beyond Context Windows of LLMs via Multi-Agent Collaboration

DGX agent

arXiv:2505.21471v2 Announce Type: replace Abstract: With the rapid advancement of post-training techniques for reasoning and information seeking, large language models (LLMs) can incorporate a large q

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Scaling Test-Time Compute for Agentic Coding

DGX agent

arXiv:2604.16529v1 Announce Type: cross Abstract: Test-time scaling has become a powerful way to improve large language models. However, existing methods are best suited to short, bounded outputs that

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

ScienceBoard: Evaluating Multimodal Autonomous Agents in Realistic Scientific Workflows

DGX agent

arXiv:2505.19897v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have extended their impact beyond Natural Language Processing, substantially fostering the development of interdi

model-releasesarxiv-cs-cl
21 Apr 2026
Local Ai

SpatialStack: Layered Geometry-Language Fusion for 3D VLM Spatial Reasoning

DGX agent

arXiv:2603.27437v2 Announce Type: replace Abstract: Large vision-language models (VLMs) still struggle with reliable 3D spatial reasoning, a core capability for embodied and physical AI systems. This

local-aiarxiv-cs-cv
21 Apr 2026
Model Releases

Spec-o3: A Tool-Augmented Vision-Language Agent for Rare Celestial Object Candidate Vetting via Automated Spectral Inspection

DGX agent

arXiv:2601.06498v2 Announce Type: replace Abstract: Due to the limited generalization and interpretability of deep learning classifiers, The final vetting of rare celestial object candidates still rel

model-releasesarxiv-cs-cl
21 Apr 2026
Local Ai

ST-pi: Structured SpatioTemporal VLA for Robotic Manipulation

DGX agent

arXiv:2604.17880v1 Announce Type: cross Abstract: Vision-language-action (VLA) models have achieved great success on general robotic tasks, but still face challenges in fine-grained spatiotemporal man

local-aiarxiv-cs-cv
21 Apr 2026
Model Releases

STaD: Scaffolded Task Design for Identifying Compositional Skill Gaps in LLMs

DGX agent

arXiv:2604.18177v1 Announce Type: new Abstract: Benchmarks are often used as a standard to understand LLM capabilities in different domains. However, aggregate benchmark scores provide limited insight

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Stop Tracking Me! Proactive Defense Against Attribute Inference Attack in LLMs

DGX agent

arXiv:2602.11528v2 Announce Type: replace-cross Abstract: Recent studies have shown that large language models (LLMs) can infer private user attributes (e.g., age, location, gender) from user-generate

researcharxiv-cs-cl
21 Apr 2026
Model Releases

Stronger Across Languages ChatGPT Images 2.0 can produce images with non-English text that’s not only rendered correctly but with language t…

DGX agent

Stronger Across Languages ChatGPT Images 2.0 can produce images with non-English text that’s not only rendered correctly but with language that flows coherently. This makes the model more globally use

model-releasesopenai--x
21 Apr 2026
Applications

Surgical Repair of Insecure Code Generation in LLMs

DGX agent

arXiv:2604.16697v1 Announce Type: cross Abstract: Large language models write production code, and yet they routinely introduce well-known vulnerabilities. We show that this is not a knowledge deficit

applicationsarxiv-cs-lg
21 Apr 2026
← Previous
1…554555556557558…1371
Next →