AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
85,202Total entries
1Added by human
85,201Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,811 results
Agents

Optimizing the Cost-Quality Tradeoff of Agentic Theorem Provers in Lean

DGX agent

arXiv:2606.04883v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used in workflows for generating formal proofs in Lean. These workflows often decompose problems into smal

agentsarxiv-cs-cl
4 Jun 2026
Tutorials
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

ParetoPilot: Zero-Surrogate Offline Multi-Objective Optimization via Infer-Perturb-Guide Diffusion

DGX agent

arXiv:2606.04468v1 Announce Type: cross Abstract: Offline multi-objective optimization (Offline MOO) aims to discover novel Pareto-optimal designs based on static datasets without expensive environmen

tutorialsarxiv-cs-ai
4 Jun 2026
Safety

Potential-Guided Flow Matching for Vision-Language-Action Policy Improvement

DGX agent

arXiv:2606.04968v1 Announce Type: new Abstract: Large vision-language-action (VLA) policies are increasingly trained as conditional generative models over action chunks. Yet deployment produces mixed-

safetyarxiv-cs-ro
4 Jun 2026
Safety

Rethinking Sales Lead Scoring with LLM-based Hierarchical Preference Ranking

DGX agent

arXiv:2606.04387v1 Announce Type: cross Abstract: Sales lead conversion in high-stakes domains (e.g., automotive, real estate) differs fundamentally from e-commerce recommendation due to prolonged dec

safetyarxiv-cs-ai
4 Jun 2026
Safety

RL Excursions during Pre-Training: Re-examining Policy Optimization for LLM training

DGX agent

arXiv:2606.04272v1 Announce Type: new Abstract: The standard LLM training pipeline applies reinforcement learning (RL) only after pre-training and supervised fine-tuning (SFT). We question this status

safetyarxiv-cs-lg
4 Jun 2026
Safety

Self-Distilled Policy Gradient

DGX agent

arXiv:2606.04036v1 Announce Type: new Abstract: On-policy self-distillation, where a language model conditions on privileged context to supervise its own generations, is a promising source of dense su

safetyarxiv-cs-lg
4 Jun 2026
Local Ai

SemBlock: Semantic Boundary Dynamic Blocks for Diffusion LLMs

DGX agent

arXiv:2606.04964v1 Announce Type: new Abstract: Diffusion language models (DLMs) generate text through iterative denoising, and blockwise decoding improves their practicality by committing tokens in l

local-aiarxiv-cs-cl
4 Jun 2026
Research

Sequential Data Poisoning in LLM Post-Training

DGX agent

arXiv:2606.04929v1 Announce Type: new Abstract: LLM post-training proceeds through multiple stages, e.g., supervised fine-tuning (SFT) followed by reinforcement learning from human feedback (RLHF) or

researcharxiv-cs-lg
4 Jun 2026
Safety

Testing Neural Networks via Bayesian-Guided Exploration of Decision Landscapes

DGX agent

arXiv:2606.04314v1 Announce Type: new Abstract: As neural networks are increasingly deployed in safety-critical domains, testing is essential to evaluate and improve their reliability. Existing testin

safetyarxiv-cs-lg
4 Jun 2026
Safety

The Right Measure for Physics-Constrained Generation: A Co-Area Correction for Posterior-Consistent PDE Inverse Problems

DGX agent

arXiv:2606.04804v1 Announce Type: new Abstract: Generative models -- diffusion and flow matching -- are increasingly used to solve partial differential equation (PDE) inverse problems, enforcing the g

safetyarxiv-cs-lg
4 Jun 2026
Local Ai

Uncertainty-Aware (Un)Supervised Few-Shot User Adaptation for On-Device Personalized Human Activity Recognition

DGX agent

arXiv:2606.04798v1 Announce Type: new Abstract: Sensor-based Human Activity Recognition (HAR) models often degrade on unseen users due to domain shifts caused by individual movement patterns and senso

local-aiarxiv-cs-lg
4 Jun 2026
Research

Uncertainty Estimation using Variance-Gated Distributions

DGX agent

arXiv:2509.08846v2 Announce Type: replace-cross Abstract: Evaluation of per-sample uncertainty quantification from neural networks is essential for decision-making involving high-risk applications. A

researcharxiv-cs-ai
4 Jun 2026
Research

UniFine: A Unified and Fine-grained Approach for Zero-shot Vision-Language Understanding

DGX agent

arXiv:2307.00862v3 Announce Type: replace-cross Abstract: Vision-language tasks, such as VQA, SNLI-VE, and VCR are challenging because they require the model's reasoning ability to understand the sema

researcharxiv-cs-cl
4 Jun 2026
Research

Unifying Information-Theoretic and Pair-Counting Clustering Similarity

DGX agent

arXiv:2511.03000v2 Announce Type: replace-cross Abstract: Comparing clusterings is central to evaluating unsupervised models, yet the many existing similarity measures can produce widely divergent, so

researcharxiv-cs-lg
4 Jun 2026
Safety

What Type of Inference is Active Inference?

DGX agent

arXiv:2606.04935v1 Announce Type: new Abstract: Active inference casts decision-making as inference, with the Expected Free Energy (EFE) unifying goal-directed and information-seeking behavior. Recent

safetyarxiv-cs-ai
4 Jun 2026
Research

You Only Train Once: Differentiable Subset Selection for Omics Data

DGX agent

arXiv:2512.17678v2 Announce Type: replace-cross Abstract: Selecting compact and informative gene subsets from single-cell transcriptomic data is essential for biomarker discovery, improving interpreta

researcharxiv-cs-ai
4 Jun 2026
Safety

A Cartesian-3j Framework for Machine Learning Interatomic Potentials

DGX agent

arXiv:2512.16882v2 Announce Type: replace-cross Abstract: Machine learning interatomic potentials (MLIPs) have brought substantial gains in the extrapolation capability in computational chemistry. How

safetyarxiv-cs-lg
3 Jun 2026
Research

A Factorized Low-Rank RNN Framework for Uncovering Independent Neural Latent Dynamics and Connectivity

DGX agent

arXiv:2511.13899v2 Announce Type: replace-cross Abstract: Low-rank recurrent neural networks (lrRNNs) are a class of models that uncover low-dimensional latent dynamics underlying neural population ac

researcharxiv-cs-lg
3 Jun 2026
Research

A Fast Screening Approach for High-dimensional Outcomes and High-dimensional Predictors

DGX agent

arXiv:2606.03018v1 Announce Type: cross Abstract: Modeling interactions among multimodal, high-dimensional data is intrinsically challenging due to ultra-high dimensionality and complex dependence str

researcharxiv-cs-lg
3 Jun 2026
Research

A Measurement-Driven Digital Twin Architecture for Plant-Level Biomass Estimation and Growth Forecasting in Hydroponic Systems

DGX agent

arXiv:2606.02796v1 Announce Type: new Abstract: Alternatives to soil-based horticulture, such as hydroponics, have been developed to respond to food distribution concerns for dense urban centers. A ne

researcharxiv-cs-ro
3 Jun 2026
Research

Act Like a Pathologist: Tissue-Aware Whole Slide Image Reasoning

DGX agent

arXiv:2603.00667v3 Announce Type: replace Abstract: Computational pathology has advanced rapidly in recent years, driven by domain-specific image encoders and growing interest in using vision-language

researcharxiv-cs-cv
3 Jun 2026
Research

Adapting Noise to Data: Generative Flows from 1D Processes

DGX agent

arXiv:2510.12636v5 Announce Type: replace-cross Abstract: The default Gaussian latent in flow-based generative models poses challenges when learning certain distributions such as heavy-tailed ones. We

researcharxiv-cs-lg
3 Jun 2026
Agents

Adaptive Latent Agentic Reasoning

DGX agent

arXiv:2606.02871v1 Announce Type: cross Abstract: Large reasoning models improve performance by generating extended chain-of-thought (CoT) reasoning, but this behavior becomes inefficient when applied

agentsarxiv-cs-ai
3 Jun 2026
Safety

Aletheia: What Makes RLVR For Code Verifiers Tick?

DGX agent

arXiv:2601.12186v3 Announce Type: replace-cross Abstract: Multi-domain thinking verifiers trained via Reinforcement Learning with Verifiable Rewards (RLVR) are a cornerstone of modern post-training. H

safetyarxiv-cs-ai
3 Jun 2026
Safety

Aligning Data-Driven Predictors with Allocation: A Decision-Focused Approach to Survival Analysis

DGX agent

arXiv:2606.02671v1 Announce Type: cross Abstract: Machine learning predictors have become essential tools for guiding automated decision making. However, a major misalignment persists: predictive mode

safetyarxiv-cs-ai
3 Jun 2026
Research

AlphaEval: A Comprehensive and Efficient Evaluation Framework for Formula Alpha Mining

DGX agent

arXiv:2508.13174v2 Announce Type: replace Abstract: Formula alpha mining, which generates predictive signals from financial data, is critical for quantitative investment. Although various algorithmic

researcharxiv-cs-ai
3 Jun 2026
Safety

ASymPO: Asymmetric-Scale Policy Optimization for Asynchronous LLM Post-Training Without Behavior Information

DGX agent

arXiv:2606.03070v1 Announce Type: cross Abstract: Asynchronous reinforcement learning can improve language-model post-training throughput by decoupling response generation from policy optimization, bu

safetyarxiv-cs-ai
3 Jun 2026
Research

BAHSD: Bridging the Long-tail Gap via Adaptive Distillation in Black-box Sequential Recommendation

DGX agent

arXiv:2606.03091v1 Announce Type: cross Abstract: Sequential recommendation systems are widely adopted but often deployed as black-box APIs, which has driven recent interest in model extraction to rep

researcharxiv-cs-ai
3 Jun 2026
Agents

Co-evolving Agent Architectures and Interpretable Reasoning for Automated Optimization

DGX agent

arXiv:2604.17708v2 Announce Type: replace Abstract: Automating operations research (OR) with large language models (LLMs) remains limited by hand-crafted reasoning--execution workflows. Complex OR tas

agentsarxiv-cs-ai
3 Jun 2026
Safety

Coherence Maximization Improves Pluralistic Alignment

DGX agent

arXiv:2606.03110v1 Announce Type: new Abstract: Aligning AI systems with diverse human values requires value specifications grounded in concrete examples, but generating such examples without extensiv

safetyarxiv-cs-cl
3 Jun 2026
Research

CORE: Conflict-Oriented Reasoning for General Multimodal Manipulation Detection

DGX agent

arXiv:2606.03066v1 Announce Type: new Abstract: The rapid rise of generative AI has made multimodal fake news increasingly realistic and pervasive, posing severe threats to public trust and social sta

researcharxiv-cs-ai
3 Jun 2026
Research

CoughSense: Five-Class Respiratory Disease Classification via Whisper Encoder Fine-Tuning and Dual-Encoder Cross-Attention Fusion with Balanced Contrastive Learning

DGX agent

arXiv:2606.02998v1 Announce Type: new Abstract: Automated cough analysis offers a path to low-cost respiratory screening, but most existing work stops at binary COVID-19 detection. A practical tool ne

researcharxiv-cs-lg
3 Jun 2026
Agents

DELTAMEM: Incremental Experience Memory for LLM Agents via Residual Trees

DGX agent

arXiv:2606.03083v1 Announce Type: new Abstract: Large Language Model (LLM)-based agents increasingly rely on memory to learn from experiences over continual interactions. However, storing experiences

agentsarxiv-cs-ai
3 Jun 2026
Safety

Denoising Tells When to Replan: Denoising-Variance Adaptive Chunking for Flow-Based Robot Policies

DGX agent

arXiv:2606.03847v1 Announce Type: new Abstract: Action chunking has become a common inference strategy for flow-based robot policies, improving action coherence by modeling multi-step temporal depende

safetyarxiv-cs-ro
3 Jun 2026
Safety

DriftSched: Adaptive QoS-Aware Scheduling under Runtime Token Drift for Multi-Tenant GPU Inference

DGX agent

arXiv:2606.02982v1 Announce Type: cross Abstract: The rapid growth of large language model (LLM) inference services has increased the demand for efficient multi-tenant GPU scheduling. While modern inf

safetyarxiv-cs-lg
3 Jun 2026
Research

DTKG: Dual-Track Knowledge Graph-Verified Reasoning Framework for Multi-Hop QA

DGX agent

arXiv:2510.16302v2 Announce Type: replace Abstract: Multi-hop reasoning for question answering (QA) plays a critical role in retrieval-augmented generation (RAG) for modern large language models (LLMs

researcharxiv-cs-ai
3 Jun 2026
Safety

Easy-to-Use Shielding for Reinforcement Learning

DGX agent

arXiv:2606.03804v1 Announce Type: new Abstract: Safe exploration is a key challenge in Reinforcement Learning (RL) that aims to prevent agents from making harmful decisions while exploring their envir

safetyarxiv-cs-lg
3 Jun 2026
Research

Efficient Transformer-Based Localized Patch Sampling for Choroid Plexus Segmentation in Multiple Sclerosis

DGX agent

arXiv:2606.03566v1 Announce Type: cross Abstract: Background: The lateral ventricle choroid plexus (LVCP) is gaining recognition as a key imaging biomarker for multiple sclerosis (MS) related to physi

researcharxiv-cs-ai
3 Jun 2026
Safety

Evaluating Transformer and LSTM Frameworks for Prediction in Ungauged Basins

DGX agent

arXiv:2606.02791v1 Announce Type: new Abstract: Watershed networks exhibit convergent topologies in which multiple tributaries merge into downstream channels,integrating diverse upstream hydrological

safetyarxiv-cs-ai
3 Jun 2026
Agents

EvoDS: Self-Evolving Autonomous Data Science Agent with Skill Learning and Context Management

DGX agent

arXiv:2606.03841v1 Announce Type: new Abstract: Recent progress in Large Language Model (LLM) agents has enabled promising advances in automated data science. However, existing approaches remain funda

agentsarxiv-cs-ai
3 Jun 2026
Research

Exploiting Verification-Generation Gap: Test-Time Reinforcement Learning with Confidence-Conditioned Verification

DGX agent

arXiv:2606.03608v1 Announce Type: cross Abstract: Test-time reinforcement learning has emerged as a promising paradigm for enhancing the complex reasoning abilities of large language models in a compl

researcharxiv-cs-ai
3 Jun 2026
Safety

Extreme Motion Generation via Hybrid Null-Space Control for Straight-Line Path Following

DGX agent

arXiv:2606.03390v1 Announce Type: new Abstract: This work studies ``extreme motion generation'', which aims to maximize the Cartesian path length along a pre-defined trajectory within the manipulator'

safetyarxiv-cs-ro
3 Jun 2026
Local Ai

extsc{CR-Seg}: Attention-Guided and CoT-Enhanced Coarse-to-Refined Reasoning Segmentation

DGX agent

arXiv:2606.03564v1 Announce Type: cross Abstract: Reasoning segmentation aims to segment target objects described by complex language through joint visual-textual reasoning. Existing methods typically

local-aiarxiv-cs-ai
3 Jun 2026
Safety

Filter, Then Reweight: Rethinking Optimization Granularity in On-Policy Distillation

DGX agent

arXiv:2606.02684v1 Announce Type: cross Abstract: On-Policy distillation (OPD) in large language models is shifting from full-trace KL supervision toward more selective training paradigms. Recent OPD

safetyarxiv-cs-ai
3 Jun 2026
Agents

FutureWeaver: Planning Test-Time Compute for Multi-Agent Systems with Modularized Collaboration

DGX agent

arXiv:2512.11213v2 Announce Type: replace Abstract: Scaling test-time computation has been shown to significantly improve large language model (LLM) performance without additional training. However, e

agentsarxiv-cs-ai
3 Jun 2026
Local Ai

Graph Regularized Non-negative Reduced Biquaternion Matrix Factorization for Color Image Recognition

DGX agent

arXiv:2606.03654v1 Announce Type: new Abstract: Non-negative reduced biquaternion matrix factorization (NRBMF) uses the product of reduced biquaternion (RB) matrices to incorporate the non-negativity

local-aiarxiv-cs-cv
3 Jun 2026
Research

Hand Trajectory Fusion for Egocentric Natural Language Query Grounding

DGX agent

arXiv:2606.02962v1 Announce Type: cross Abstract: Egocentric Natural Language Query (NLQ) grounding asks a model to localize, in a long first-person video, the temporal interval that answers a free-fo

researcharxiv-cs-ai
3 Jun 2026
Agents

Handoff Debt: The Rediscovery Cost When Coding Agents Take Over Interrupted Tasks

DGX agent

arXiv:2606.02875v1 Announce Type: new Abstract: Coding-agent benchmarks evaluate whether a single uninterrupted agent can resolve a repository issue. Real software work is messier: tasks are interrupt

agentsarxiv-cs-ai
3 Jun 2026
← Previous
1…764765766767768…1038
Next →