AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

88,419Total entries
1Added by human
88,418Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,638 results
3 Jun 2026

SLU-2K: A Question-Based Benchmark for Semantic Evaluation of Sign Language Translation

Model ReleasesDGX agent

arXiv:2606.03788v1 Announce Type: new Abstract: Sign Language Translation (SLT) is typically evaluated with surface-form metrics such as BLEU and ROUGE, which reward lexical overlap but do not directl

StepFinder: A Temporal Semantic Framework for Failure Attribution in Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2606.03467v1 Announce Type: new Abstract: LLM-based multi-agent systems exhibit remarkable collaborative capabilities in complex multi-step tasks. However, these systems are highly sensitive to

Synthesize and Reward -- Reinforcement Learning for Multi-Step Tool Use in Live Environments

AgentsDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.03892v1 Announce Type: cross Abstract: Training LLMs to orchestrate multi-step tool calls is held back by three coupled obstacles: realistic stateful execution environments are costly to bu

TadA-Bench: A Million-Variant Benchmark for Future-Round Discovery Toward Agentic Protein Engineering

Model ReleasesDGX agent

arXiv:2606.02624v1 Announce Type: cross Abstract: AI for scientific discovery is entering an agentic era, where protein-engineering systems are expected to prioritize future wet-lab experiments rather

The DeepSpeak-Agentic Dataset

Model ReleasesDGX agent

arXiv:2606.03686v1 Announce Type: new Abstract: We present DeepSpeak-Agentic, a dataset of videos comprising over 37 hours of semi-structured conversations between a human and an embodied AI agent. We

The Efficiency vs. Accuracy Trade-off: Optimizing RAG-Enhanced LLM Recommender Systems Using Multi-Head Early Exit

ResearchDGX agent

arXiv:2501.02173v2 Announce Type: replace-cross Abstract: The deployment of Large Language Models (LLMs) in recommender systems for predicting Click-Through Rates (CTR) necessitates a delicate balance

Towards Characterizing Scientific Image Utility and Upgradability

Model ReleasesDGX agent

arXiv:2606.03401v1 Announce Type: new Abstract: Scientific images function as critical evidence in research communication, yet their integrity faces unprecedented threats from AI-generated content tha

Towards Fair Graph Prompting: A Dual-Prompt Mechanism for Mitigating Attribute and Structural Bias

Model ReleasesDGX agent

arXiv:2510.23469v2 Announce Type: replace Abstract: Self-supervised pre-training on unlabeled graph data has become a common paradigm for Graph Neural Networks (GNNs). However, an objective gap often

Training-Free Multi-Concept LoRA Composition with Prompt-Aware Weighting

ApplicationsDGX agent

arXiv:2606.03792v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) successfully enables personalization in text-to-image generation by adapting pre-trained diffusion models to specific visual

TSQAgent: Rating Time Series Data Quality via Dedicated Agentic Reasoning

Model ReleasesDGX agent

arXiv:2606.03629v1 Announce Type: new Abstract: Assessing the quality of time series (TS) data is fundamental yet inherently challenging due to the multifaceted nature of quality dimensions. Recently,

What’s new in serverless Managed Service for Apache Spark

HardwareDGX agent

Whether you use it for data preparation, real-time interactive queries, AI model training, or something entirely different, running Apache Spark at scale is demanding — you shouldn’t have to manage th

WRIT: Write-Read Intensive Trajectory Synthesis for Multi-Turn User-Facing Agents

Model ReleasesDGX agent

arXiv:2606.02908v1 Announce Type: cross Abstract: Multi-turn user-facing agents must infer user intent from incomplete requests, collect missing information through dialogue and tools, and execute val

2 Jun 2026

A Comparative Analysis of Machine Learning Algorithms for Multi-Task Prediction of the Parameters of the Pectin Hydrolysis--Extraction Process

Model ReleasesDGX agent

arXiv:2606.00821v1 Announce Type: new Abstract: This study addresses the challenge of controlling a complex, multi-parameter technological process -- pectin hydrolysis--extraction -- using machine lea

A Shared Valence Axis Across Modern LLMs and Human EEG: The Saturation Regularity

SafetyDGX agent

arXiv:2606.00129v1 Announce Type: cross Abstract: Large language models (LLMs) have emerged as powerful representation learners whose internal features increasingly align with human cognition. We stud

Absorbing Complexity: An Interaction-Native Knowledge Harness for Financial LLM Agents

Model ReleasesDGX agent

arXiv:2606.01886v1 Announce Type: new Abstract: Financial AI agents often fail for a simple reason: they make users carry the complexity. A user must repeatedly restate goals, risk preferences, portfo

Adaptive Order Policies for Masked Diffusion

SafetyDGX agent

arXiv:2606.00295v1 Announce Type: new Abstract: Masked diffusion models have seen great success in capturing data distributions over discrete sequences in domains such as text and proteins. These mode

Adaptive Time Series Reasoning via Segment Selection

SafetyDGX agent

arXiv:2602.18645v2 Announce Type: replace Abstract: Time series reasoning tasks often start with a natural language question and require targeted analysis of a time series. Evidence may span the full

Agent-R1: A Unified and Modular Framework for Agentic Reinforcement Learning

AgentsDGX agent

arXiv:2511.14460v2 Announce Type: replace Abstract: Large language models (LLMs) have rapidly evolved from single-turn text generators into the foundation of increasingly capable agents. As these agen

Agent Tools Orchestration Leaks More: Dataset, Benchmark, and Mitigation

Model ReleasesDGX agent

arXiv:2512.16310v3 Announce Type: replace-cross Abstract: LLM-based agents increasingly use multiple external tools to complete complex tasks. We study Tools Orchestration Privacy Risk (TOP-R): an age

AgentDS Technical Report: Benchmarking the Future of Human-AI Collaboration in Domain-Specific Data Science

Model ReleasesDGX agent

arXiv:2603.19005v2 Announce Type: replace-cross Abstract: Data science plays a critical role in transforming complex data into actionable insights across numerous domains. Recent developments in large

{alpha}Depth: Learning Single-Pass Soft Boundary Decomposition for Stereo Conversion

Local AiDGX agent

arXiv:2606.00386v1 Announce Type: new Abstract: Accurately modeling soft boundaries, e.g., hair and defocus blur, is a fundamental challenge in stereo conversion due to the ambiguous blending of foreg

Amortized Predictability-aware Training Framework for Time Series Forecasting and Classification

Local AiDGX agent

arXiv:2602.16224v2 Announce Type: replace Abstract: Time series data are prone to noise in various domains, and training samples may contain low-predictability patterns that deviate from the normal da

Approximating f-Divergences with Rank Statistics

Model ReleasesDGX agent

arXiv:2601.22784v2 Announce Type: replace-cross Abstract: We introduce a rank-statistic approximation of f-divergences that avoids explicit density-ratio estimation by working directly with the distri

ASAP: Advancing Medical Volumetric Representation Learning with Anatomy-aware Semantically-adaptive Pre-training

Model ReleasesDGX agent

arXiv:2606.00602v1 Announce Type: new Abstract: Learning transferable and interpretable representations from medical volumetric scans remains challenging due to complex anatomical structures and weak,

Auteur: Language-Driven Cinematographic Framing for Human-Centric Video Generation

ApplicationsDGX agent

arXiv:2606.01900v1 Announce Type: new Abstract: Generative video models have achieved remarkable visual fidelity and temporal coherence, yet intentional camera control remains elusive. Existing framew

Autopilot-Preserving Residual Q-Learning with HJB-Inspired Finite-Action Risk Filtering for Fixed-Wing UAV Command Supervision

Model ReleasesDGX agent

arXiv:2606.01397v1 Announce Type: cross Abstract: A fixed-wing UAV must hold airspeed, altitude, and heading references under wind, gusts, and turbulence, channels coupled so that correcting one can d

banger writeup on what it takes to make MiniMax M3 go brrrr!

ToolsDGX agent

This post likely provides technical insights and optimization strategies for running or fine-tuning the MiniMax M3 model efficiently, discussing performance enhancements and practical implementation t

Bayesian Inference of Nonlinear Malaria Dynamics in Ghana via an Ensemble Markov Chain Monte Carlo Sampler

Model ReleasesDGX agent

arXiv:2606.00783v1 Announce Type: cross Abstract: Reliable quantification of malaria dynamics in sub-Saharan Africa is hindered by short, noisy, and spatially heterogeneous surveillance records. In Gh

Beyond Independent Manipulation: Individual Fairness-aware Strategic Classification with Peer Imitation

SafetyDGX agent

arXiv:2606.00827v1 Announce Type: cross Abstract: Strategic classification (SC) investigates scenarios where agents manipulate their features to obtain favorable decisions from predictive models. Exis

Beyond Task Success: Behavioral and Representational Diagnostics for WAM and VLA

ResearchDGX agent

arXiv:2606.01095v1 Announce Type: cross Abstract: Vision-language-action (VLA) policies and World-Action Models (WAM) represent two increasingly important paradigms for robotic manipulation. However,

Beyond Visual Memory: Mechanistic Diagnostics of Latent Visual Reasoning

ResearchDGX agent

arXiv:2606.01287v1 Announce Type: cross Abstract: Recent latent visual reasoning methods achieve substantial gains by inserting continuous latent tokens into multimodal language models. These gains ar

BitsMoE: Efficient Spectral Energy-Guided Bit Allocation for MoE LLM Quantization

ResearchDGX agent

arXiv:2606.00079v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) large language models reduce per-token computation through sparse expert activation, but their deployment remains memory-inte

Boundary-Protection W8A8 HiFloat8 Quantization for Large-Scale Text-to-Video Diffusion Transformers

Model ReleasesDGX agent

arXiv:2606.00957v1 Announce Type: new Abstract: We present a post-training quantization (PTQ) approach for Wan2.1-T2V-14B, a 14-billion-parameter text-to-video diffusion transformer, targeting the W8A

Bridging the Knowledge-Prediction Gap in LLMs on Multiple-Choice Questions

ResearchDGX agent

arXiv:2509.23782v4 Announce Type: replace Abstract: While large language models (LLMs) perform strongly on diverse tasks, their trustworthiness is limited by erratic behavior that is unfaithful to the

CAFOSat: A Strongly Annotated Dataset for Infrastructure-Aware CAFO Mapping Using High-Resolution Imagery

Model ReleasesDGX agent

arXiv:2606.00548v1 Announce Type: cross Abstract: Concentrated Animal Feeding Operations (CAFOs) play an important role in agricultural production but are also associated with environmental, public he

Can we design legal agent verifiers that are up to 1,000x cheaper? Verifiers are LLM judges that check an agent’s work against rubric criter…

Model ReleasesDGX agent

Can we design legal agent verifiers that are up to 1,000x cheaper? Verifiers are LLM judges that check an agent’s work against rubric criteria: they're used both in agent benchmarking and as reward si

Capability Self-Assessment: Teaching LLMs to Know Their Limits

Local AiDGX agent

arXiv:2606.00251v1 Announce Type: new Abstract: The ability to recognize one's own limitations and decide whether to solve a problem or delegate is fundamental for reliable intelligent systems. Yet we

Certificates without Electrons? Theory and Evidence on Impacts from AI-Driven Power Demand

Local AiDGX agent

arXiv:2606.00811v1 Announce Type: cross Abstract: Data centers now account for 4.4% of United States electricity demand, yet the grid-level effectiveness of the renewable energy certificates (RECs) an

CLAW: A Vision-Language-Action Framework for Weight-Aware Robotic Grasping

SafetyDGX agent

arXiv:2509.14143v2 Announce Type: replace Abstract: Vision-language-action (VLA) models have recently emerged as a promising paradigm for robotic control, enabling end-to-end policies that ground natu

Co-training with Ego-centric Video and Demonstration for Robot Navigation Task

ResearchDGX agent

arXiv:2606.01951v1 Announce Type: new Abstract: Vision-language-action (VLA) models are promising for diverse robotic tasks, but their performance heavily depends on large-scale high-quality training

Concept Heterogeneity-aware Representation Steering

ResearchDGX agent

arXiv:2603.02237v2 Announce Type: replace-cross Abstract: Representation steering offers a lightweight mechanism for controlling the behavior of large language models (LLMs) by intervening on internal

Continual Learning involves engineering whole systems including Efficient Verifiers to make RL/fine-tuning and running evaluations much chea…

Model ReleasesDGX agent

Continual Learning involves engineering whole systems including Efficient Verifiers to make RL/fine-tuning and running evaluations much cheaper at scale! some initial work we’re releasing from LangCha

CRAB-Bench: Evaluating LLM Agents under Complex Task Dependencies and Human-aligned User Simulation

Model ReleasesDGX agent

arXiv:2606.01815v1 Announce Type: new Abstract: Evaluating LLM agents in realistic service scenarios requires complex task dependencies, imperfect user behavior, and an evaluation that accommodates mu

CRAFT: Fine-Grained Cost-Aware Expert Replication For Efficient Mixture-of-Experts Serving

HardwareDGX agent

arXiv:2603.28768v2 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) has recently emerged as the mainstream architecture for efficiently scaling large language models while maintaining n

CURP: Codebook-based Continuous User Representation for Personalized Generation with LLMs

ResearchDGX agent

arXiv:2602.00742v2 Announce Type: replace Abstract: User modeling characterizes individuals through their preferences and behavioral patterns to enable personalized simulation and generation with Larg

Design-MLLM: A Reinforcement Alignment Framework for Verifiable and Aesthetic Interior Design

Model ReleasesDGX agent

arXiv:2603.13312v2 Announce Type: replace-cross Abstract: Interior design is a requirements-to-visual-plan generation process that must simultaneously satisfy verifiable spatial feasibility and compar

EEG-FuseFormer: A Transformer-Driven Feature Fusion Framework for Seizure Onset Prediction

ResearchDGX agent

arXiv:2606.02166v1 Announce Type: new Abstract: Epilepsy is one of the most common neurological disorders globally, characterized by recurring seizures and significantly impacting the quality of life.

Ego-METAS: Egocentric online Multimodal Energy-efficient Temporal Action Segmentation benchmark

Model ReleasesDGX agent

arXiv:2606.02246v1 Announce Type: new Abstract: To operate in the physical world, embodied agents must perceive their environment in an 'always-on' fashion, selectively accessing the most informative

Enhancing Blind Source Separation with Dissociative Principal Component Analysis

Model ReleasesDGX agent

arXiv:2411.12321v2 Announce Type: replace Abstract: Principal component analysis (PCA) and its sparse variants (sPCA) are widely used as a precursor to independent component analysis (ICA) for blind s

Evaluating Reliability Asymmetries in Chinese Factual Search and AI Answers

Model ReleasesDGX agent

arXiv:2602.22221v2 Announce Type: replace-cross Abstract: Search engines and AI-powered systems increasingly mediate access to factual information, yet their reliability remains difficult to evaluate

Fair Finetuning Mitigates Distribution Inference Attacks

SafetyDGX agent

arXiv:2606.01719v1 Announce Type: cross Abstract: Machine learning models trained on sensitive data can inadvertently leak population-level information about their training distributions -- a threat k

Fast Generalization after Interpolation via Critically Damped Momentum Optimization

Local AiDGX agent

arXiv:2606.01521v1 Announce Type: new Abstract: A central problem in machine learning is that models can achieve near-perfect training performance while generalizing substantially less well to unseen

Faster Synchronous On-Policy RL via Straggler-Aware Group Sizing

Model ReleasesDGX agent

arXiv:2606.02218v1 Announce Type: cross Abstract: Synchronous reinforcement learning methods such as Group Relative Policy Optimization (GRPO) provide stable and reproducible on-policy training, but t

FastSLM: Hierarchical Temporal Abstraction for Efficient Long-Form Speech Adaptation

ResearchDGX agent

arXiv:2601.06199v3 Announce Type: replace-cross Abstract: Scaling Multimodal Large Language Models (MLLMs) to long-form speech is bottlenecked by the explosive growth of input tokens. Unlike images or

Federated Learning via Variational Bayesian Inference: Personalization, Sparsity and Clustering

ResearchDGX agent

arXiv:2303.04345v2 Announce Type: replace Abstract: Federated learning (FL) is a promising framework that models distributed machine learning while protecting the privacy of clients. However, FL suffe

FedS2R: One-Shot Federated Domain Generalization for Synthetic-to-Real Semantic Segmentation in Autonomous Driving

AgentsDGX agent

arXiv:2507.19881v2 Announce Type: replace-cross Abstract: Federated domain generalization has shown promising progress in image classification by enabling collaborative training across multiple client

FLAME: Physics-Guided Neural Operators for Onboard Satellite Methane Detection in Hyperspectral Imagery

Model ReleasesDGX agent

arXiv:2606.01577v1 Announce Type: new Abstract: Methane is a major driver of near-term climate change, and rapidly identifying its emission sources is a critical climate intervention. Spaceborne hyper

Generative AI and Digital Ecosystem Resilience: A Proactive Lifecycle-Based Survey

AgentsDGX agent

arXiv:2606.00136v1 Announce Type: cross Abstract: The proliferation of adversarial synthetic content, accelerated by Generative AI (GenAI) is rendering traditional reactive detection methods ineffecti

Geometric Erasure by Contrastive Velocity Matching in Rectified Flows

ResearchDGX agent

arXiv:2606.00140v1 Announce Type: cross Abstract: While the rapid adoption of multimodal generative models offers immense potential, it has also increased the risks of harmful content synthesis, deepf

GREAT: Generalizable Backdoor Attacks in RLHF via Emotion-Aware Trigger Synthesis

Model ReleasesDGX agent

arXiv:2510.09260v2 Announce Type: replace-cross Abstract: Recent work has shown that RLHF is highly susceptible to backdoor attacks. However, existing methods often rely on rare tokens or fixed trigge

← Previous
1…541542543544545…1061
Next →