AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,376
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,170
  • Local Ai4,930
  • Model Releases23,883
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,376
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,170
  • Local Ai4,930
  • Model Releases23,883
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

88,376Total entries
1Added by human
88,375Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,602 results
21 Apr 2026

Conformal Prediction-Based MPC for Stochastic Linear Systems

ResearchDGX agent

arXiv:2512.10738v2 Announce Type: replace-cross Abstract: We propose a stochastic model predictive control (MPC) framework for linear systems subject to joint-in-time chance constraints under unknown

ConforNets: Latents-Based Conformational Control in OpenFold3

ResearchDGX agent

arXiv:2604.18559v1 Announce Type: cross Abstract: Models from the AlphaFold (AF) family reliably predict one dominant conformation for most well-ordered proteins but struggle to capture biologically r

Continual Safety Alignment via Gradient-Based Sample Selection

SafetyDGX agent

arXiv:2604.17215v1 Announce Type: new Abstract: Large language models require continuous adaptation to new tasks while preserving safety alignment. However, fine-tuning on even benign data often compr

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Countdown-Code: A Testbed for Studying The Emergence and Generalization of Reward Hacking in RLVR

ResearchDGX agent

arXiv:2603.07084v2 Announce Type: replace-cross Abstract: Reward hacking is a form of misalignment in which models overoptimize proxy rewards without genuinely solving the underlying task. Precisely m

Cross-Modal Bayesian Low-Rank Adaptation for Uncertainty-Aware Multimodal Learning

Model ReleasesDGX agent

arXiv:2604.16657v1 Announce Type: new Abstract: Large pre-trained language models are increasingly adapted to downstream tasks using parameter-efficient fine-tuning (PEFT), but existing PEFT methods a

Decomposing the Depth Profile of Fine-Tuning

ResearchDGX agent

arXiv:2604.17177v1 Announce Type: new Abstract: Fine-tuning adapts pretrained networks to new objectives. Whether the resulting depth profile of representational change reflects an intrinsic property

DEM Refinement and Validation on the Lunar Surface Using Shape-from-Shading with Chandrayaan-2 OHRC Imagery

Model ReleasesDGX agent

arXiv:2604.17436v1 Announce Type: new Abstract: This study presents a Shape from Shading (SfS) framework to enhance sub-metre resolution lunar digital elevation models (DEMs) using imagery from the Or

Depth Registers Unlock W4A4 on SwiGLU: A Reader/Generator Decomposition

Model ReleasesDGX agent

arXiv:2604.18128v1 Announce Type: new Abstract: We study post-training W4A4 quantization in a controlled 300M-parameter SwiGLU decoder-only language model trained on 5B tokens of FineWeb-Edu, and ask

Diffusion-Based Optimization for Accelerated Convergence of Redundant Dual-Arm Minimum Time Problems

ResearchDGX agent

arXiv:2604.16670v1 Announce Type: new Abstract: We present a framework leveraging a novel variant of the model-based diffusion algorithm to minimize the time required for a redundant dual-arm robot co

Domain-oriented RAG Assessment (DoRA): Synthetic Benchmarking for RAG-based Question Answering on Defense Documents

Model ReleasesDGX agent

arXiv:2604.17943v1 Announce Type: new Abstract: Open-domain RAG benchmarks over public corpora can overestimate deployment performance due to pretraining overlap and weak attribution requirements. We

DREAM: Dynamic Retinal Enhancement with Adaptive Multi-modal Fusion for Expert Precision Medical Report Generation

Model ReleasesDGX agent

arXiv:2604.17209v1 Announce Type: new Abstract: Automating medical reports for retinal images requires a sophisticated blend of visual pattern recognition and deep clinical knowledge. Current Large Vi

Driving in Corner Case: A Real-World Adversarial Closed-Loop Evaluation Platform for End-to-End Autonomous Driving

SafetyDGX agent

arXiv:2512.16055v2 Announce Type: replace Abstract: Safety-critical corner cases, difficult to collect in the real world, are crucial for evaluating end-to-end autonomous driving. Adversarial interact

Dual-stream Spatio-Temporal GCN-Transformer Network for 3D Human Pose Estimation

Model ReleasesDGX agent

arXiv:2604.17688v1 Announce Type: new Abstract: 3D human pose estimation is a classic and important research direction in the field of computer vision. In recent years, Transformer-based methods have

DualToken: Towards Unifying Visual Understanding and Generation with Dual Visual Vocabularies

ResearchDGX agent

arXiv:2503.14324v3 Announce Type: replace-cross Abstract: The differing representation spaces required for visual understanding and generation pose a challenge in unifying them within the autoregressi

DuConTE: Dual-Granularity Text Encoder with Topology-Constrained Attention for Text-attributed Graphs

Model ReleasesDGX agent

arXiv:2604.17411v1 Announce Type: new Abstract: Text-attributed graphs integrate semantic information of node texts with topological structure, offering significant value in various applications such

EasyVideoR1: Easier RL for Video Understanding

Model ReleasesDGX agent

arXiv:2604.16893v1 Announce Type: new Abstract: Reinforcement learning from verifiable rewards (RLVR) has demonstrated remarkable effectiveness in improving the reasoning capabilities of large languag

End-to-end Listen, Look, Speak and Act

Model ReleasesDGX agent

arXiv:2510.16756v2 Announce Type: replace-cross Abstract: Human interaction is inherently multimodal and full-duplex: we listen while watching, speak while acting, and fluidly adapt to turn-taking and

Enhancing Zero-shot Personalized Image Aesthetics Assessment with Profile-aware Multimodal LLM

ResearchDGX agent

arXiv:2604.17233v1 Announce Type: new Abstract: Personalized image aesthetics assessment (PIAA) aims to predict an individual user's subjective rating of an image, which requires modeling user-specifi

Error as Signal: Stiffness-Aware Diffusion Sampling via Embedded Runge-Kutta Guidance

Model ReleasesDGX agent

arXiv:2603.03692v2 Announce Type: replace Abstract: Classifier-Free Guidance (CFG) has established the foundation for guidance mechanisms in diffusion models, showing that well-designed guidance proxi

Explanation Bias is a Product: Revealing the Hidden Lexical and Position Preferences in Post-Hoc Feature Attribution

SafetyDGX agent

arXiv:2512.11108v3 Announce Type: replace Abstract: Good quality explanations strengthen the understanding of language models and data. Feature attribution methods, such as Integrated Gradient, are a

FairLogue: Evaluating Intersectional Fairness across Clinical Machine Learning Use Cases using the All of Us Research Program

SafetyDGX agent

arXiv:2604.16450v1 Announce Type: cross Abstract: Intersectional biases in healthcare data can produce compound disparities in clinical machine learning models, yet most fairness evaluations assess de

Faithfulness vs. Safety: Evaluating LLM Behavior Under Counterfactual Medical Evidence

SafetyDGX agent

arXiv:2601.11886v2 Announce Type: replace Abstract: In high-stakes domains like medicine, it may be generally desirable for models to faithfully adhere to the context provided. But what happens if the

FlashFPS: Efficient Farthest Point Sampling for Large-Scale Point Clouds via Pruning and Caching

Model ReleasesDGX agent

arXiv:2604.17720v1 Announce Type: cross Abstract: Point-based Neural Networks (PNNs) have become a key approach for point cloud processing. However, a core operation in these models, Farthest Point Sa

FrameDiT: Diffusion Transformer with Matrix Attention for Efficient Video Generation

ResearchDGX agent

arXiv:2603.09721v2 Announce Type: replace Abstract: High-fidelity video generation remains challenging for diffusion models due to the difficulty of modeling complex spatio-temporal dynamics efficient

Frequency-guided Multi-level Reasoning for Scene Graph Generation in Video

ResearchDGX agent

arXiv:2604.17298v1 Announce Type: new Abstract: Video Scene Graph Generation aims to obtain structured semantic representations of objects and their relationships in videos for high-level understandin

From 2:4 to 8:16 sparsity patterns in LLMs for Outliers and Weights with Variance Correction

ResearchDGX agent

arXiv:2507.03052v2 Announce Type: replace Abstract: As large language models (LLMs) grow in size, efficient compression techniques like quantization and sparsification are critical. While quantization

GeoRC: A Benchmark for Geolocation Reasoning Chains

Model ReleasesDGX agent

arXiv:2601.21278v2 Announce Type: replace-cross Abstract: Vision Language Models (VLMs) are good at recognizing the global location of a photograph -- their geolocation prediction accuracy rivals the

Harness Engineering Without the Hype: 🦄 AI That Works #54 https://x.com/i/broadcasts/1DxLdvQrDwkxm

AgentsDGX agent

This episode of 'AI That Works' discusses practical approaches to harness engineering in AI systems, likely focusing on techniques for effectively prompting and controlling AI model behavior beyond ma

High-Resolution Visual Reasoning via Multi-Turn Grounding-Based Reinforcement Learning

SafetyDGX agent

arXiv:2507.05920v2 Announce Type: replace Abstract: State-of-the-art large multi-modal models (LMMs) face challenges when processing high-resolution images, as these inputs are converted into enormous

HopWeaver: Cross-Document Synthesis of High-Quality and Authentic Multi-Hop Questions

ResearchDGX agent

arXiv:2505.15087v3 Announce Type: replace Abstract: Multi-Hop Question Answering (MHQA) is crucial for evaluating the model's capability to integrate information from diverse sources. However, creatin

How to Ground a Korean AI Agent in Real Demographics with Synthetic Personas

AgentsDGX agent

This article explains how to ground Korean AI agents in realistic demographic data by utilizing synthetic personas, likely leveraging NVIDIA's Nemotron models available on Hugging Face. The approach e

HUGGING FACE JUST AUTOMATED THEIR ENTIRE POST-TRAINING TEAM WITH AN AGENT. It reads papers, runs GPU experiments, iterates, and builds resea…

Model ReleasesDGX agent

HUGGING FACE JUST AUTOMATED THEIR ENTIRE POST-TRAINING TEAM WITH AN AGENT. It reads papers, runs GPU experiments, iterates, and builds research-backed models autonomously. Pushed a benchmark from 10%

IMA-MoE: An Interpretable Modality-Aware Mixture-of-Experts Framework for Characterizing the Neurobiological Signatures of Binge Eating Disorder

ResearchDGX agent

arXiv:2604.17028v1 Announce Type: new Abstract: Binge eating disorder (BED) is the most prevalent eating disorder. However, current diagnostic frameworks remain largely grounded in symptom-based crite

Implicit neural representations as a coordinate-based framework for continuous environmental field reconstruction from sparse ecological observations

SafetyDGX agent

arXiv:2604.18083v1 Announce Type: new Abstract: Reconstructing continuous environmental fields from sparse and irregular observations remains a central challenge in environmental modelling and biodive

Inflated Excellence or True Performance? Rethinking Medical Diagnostic Benchmarks with Dynamic Evaluation

Model ReleasesDGX agent

arXiv:2510.09275v2 Announce Type: replace Abstract: Medical diagnostics is a high-stakes and complex domain that is critical to patient care. However, current evaluations of large language models (LLM

Kimi K2.6 is now live inside Anything!

Model ReleasesDGX agent

Kimi K2.6, an AI model from Moonshot, has been integrated into the Anything platform. This update likely enables users to access Kimi's capabilities directly within the Anything application interface.

LayerCache: Exploiting Layer-wise Velocity Heterogeneity for Efficient Flow Matching Inference

Model ReleasesDGX agent

arXiv:2604.16492v1 Announce Type: new Abstract: Flow Matching models achieve state-of-the-art image generation quality but incur substantial inference cost due to iterative denoising through large Tra

Learning Stable Predictors from Weak Supervision under Distribution Shift

Model ReleasesDGX agent

arXiv:2604.05002v2 Announce Type: replace Abstract: Learning from weak, proxy, or relative supervision is common when ground-truth labels are unavailable, but robustness under distribution shift remai

Learning to Control Summaries with Score Ranking

Model ReleasesDGX agent

arXiv:2604.17197v1 Announce Type: new Abstract: Recent advances in summarization research focus on improving summary quality across multiple criteria, such as completeness, conciseness, and faithfulne

Learning to Correct: Calibrated Reinforcement Learning for Multi-Attempt Chain-of-Thought

ResearchDGX agent

arXiv:2604.17912v1 Announce Type: new Abstract: State-of-the-art reasoning models utilize long chain-of-thought (CoT) to solve increasingly complex problems using more test-time computation. In this w

Learning to Retrieve User History and Generate User Profiles for Personalized Persuasiveness Prediction

Model ReleasesDGX agent

arXiv:2601.05654v3 Announce Type: replace Abstract: Estimating the persuasiveness of messages is critical in various applications, from recommender systems to safety assessment of LLMs. While it is im

Lil: Less is Less When Applying Post-Training Sparse-Attention Algorithms in Long-Decode Stage

ResearchDGX agent

arXiv:2601.03043v3 Announce Type: replace Abstract: Large language models (LLMs) demonstrate strong capabilities across a wide range of complex tasks and are increasingly deployed at scale, placing si

LIVE: Leveraging Image Manipulation Priors for Instruction-based Video Editing

Model ReleasesDGX agent

arXiv:2604.17021v1 Announce Type: new Abstract: Video editing aims to modify input videos according to user intent. Recently, end-to-end training methods have garnered widespread attention, constructi

LLM Hypnosis: Exploiting User Feedback for Unauthorized Knowledge Injection to All Users

ResearchDGX agent

arXiv:2507.02850v3 Announce Type: replace Abstract: We describe a vulnerability in language models (LMs) trained with user feedback, whereby a single user can persistently alter LM knowledge and behav

Lorentz Framework for Semantic Segmentation

TutorialsDGX agent

arXiv:2604.16836v1 Announce Type: new Abstract: Semantic segmentation in hyperbolic space enables compact modeling of hierarchical structure while providing inherent uncertainty quantification. Prior

ltzGLUE: Luxembourgish General Language Understanding Evaluation

Model ReleasesDGX agent

arXiv:2604.17976v1 Announce Type: new Abstract: This paper presents ltzGLUE, the first Natural Language Understanding (NLU) benchmark for Luxembourgish (LTZ) based on the popular GLUE benchmark for En

MambaKick: Early Penalty Direction Prediction from HAR Embeddings

ApplicationsDGX agent

arXiv:2604.16588v1 Announce Type: new Abstract: Penalty kicks in soccer are decided under extreme time constraints, where goalkeepers benefit from anticipating shot direction from the kickers motion b

MemBuilder: Reinforcing LLMs for Long-Term Memory Construction via Attributed Dense Rewards

Model ReleasesDGX agent

arXiv:2601.05488v3 Announce Type: replace Abstract: Maintaining consistency in long-term dialogues remains a fundamental challenge for LLMs, as standard retrieval mechanisms often fail to capture the

Multimodal In-context Learning for ASR of Low-resource Languages

Model ReleasesDGX agent

arXiv:2601.05707v2 Announce Type: replace Abstract: Automatic speech recognition (ASR) still covers only a small fraction of the world's languages, mainly due to supervised data scarcity. In-context l

MUSEG: Reinforcing Video Temporal Understanding via Timestamp-Aware Multi-Segment Grounding

ResearchDGX agent

arXiv:2505.20715v2 Announce Type: replace-cross Abstract: Video temporal understanding is crucial for multimodal large language models (MLLMs) to reason over events in videos. Despite recent advances

Network-wide Freeway Traffic Estimation Using Sparse Sensor Data: A Dirichlet Graph Auto-Encoder Approach

Local AiDGX agent

arXiv:2503.15845v2 Announce Type: replace Abstract: Network-wide Traffic State Estimation (TSE), which aims to infer a complete image of network traffic states with sparsely deployed sensors, plays a

One-Step Diffusion with Inverse Residual Fields for Unsupervised Industrial Anomaly Detection

ResearchDGX agent

arXiv:2604.18393v1 Announce Type: new Abstract: Diffusion models have achieved outstanding performance in unsupervised industrial anomaly detection (uIAD) by learning a manifold of normal data under t

OPeRA: A Dataset of Observation, Persona, Rationale, and Action for Evaluating LLMs on Human Online Shopping Behavior Simulation

Model ReleasesDGX agent

arXiv:2506.05606v5 Announce Type: replace Abstract: Can large language models (LLMs) accurately simulate the next web action of a specific user? While LLMs have shown promising capabilities in generat

Privacy-R1: Privacy-Aware Multi-LLM Agent Collaboration via Reinforcement Learning

Local AiDGX agent

arXiv:2510.16054v2 Announce Type: replace-cross Abstract: When users submit queries to Large Language Models (LLMs), their prompts can often contain sensitive data, forcing a difficult choice: Send th

Progressive Online Video Understanding with Evidence-Aligned Timing and Transparent Decisions

Model ReleasesDGX agent

arXiv:2604.18459v1 Announce Type: new Abstract: Visual agents operating in the wild must respond to queries precisely when sufficient evidence first appears in a video stream, a critical capability th

Pseudo2Real: Task Arithmetic for Pseudo-Label Correction in Automatic Speech Recognition

Model ReleasesDGX agent

arXiv:2510.08047v2 Announce Type: replace-cross Abstract: Robust ASR under domain shift is crucial because real-world systems encounter unseen accents and domains with limited labeled data. Although p

ReconVLA: An Uncertainty-Guided and Failure-Aware Vision-Language-Action Framework for Robotic Control

ApplicationsDGX agent

arXiv:2604.16677v1 Announce Type: new Abstract: Vision-language-action (VLA) models have emerged as generalist robotic controllers capable of mapping visual observations and natural language instructi

ResearchBench: Benchmarking LLMs in Scientific Discovery via Inspiration-Based Task Decomposition

Model ReleasesDGX agent

arXiv:2503.21248v3 Announce Type: replace Abstract: Large language models (LLMs) have shown potential in assisting scientific research, yet their ability to discover high-quality research hypotheses r

Revisiting Active Sequential Prediction-Powered Mean Estimation

Model ReleasesDGX agent

arXiv:2604.18569v1 Announce Type: cross Abstract: In this work, we revisit the problem of active sequential prediction-powered mean estimation, where at each round one must decide the query probabilit

Sampling for Quality: Training-Free Reward-Guided LLM Decoding via Sequential Monte Carlo

ResearchDGX agent

arXiv:2604.16453v1 Announce Type: new Abstract: We introduce a principled probabilistic framework for reward-guided decoding in large language models, addressing the limitations of standard decoding m

← Previous
1…427428429430431…1061
Next →