AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
AllBlog
90,914Total entries
1Added by human
90,913Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,676 results
Safety

Reliability-Aware LLM Alignment from Inconsistent Human Feedback

DGX agent

arXiv:2607.20515v1 Announce Type: new Abstract: Reinforcement Learning from Human Feedback (RLHF) is critical for aligning Large Language Models (LLMs) with human preferences. However, its efficacy is

safetyarxiv-cs-ai
24 Jul 2026
Model Releases

Rushes: A Human Preference Dataset for Pluralistic Alignment

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2607.20767v1 Announce Type: new Abstract: We introduce Rushes, a dataset and benchmark for studying revealed human engagement preferences in interactive narrative environments. Rushes is collect

model-releasesarxiv-cs-cl
24 Jul 2026
Safety

SalesLoop: Reinforcement Learning from Performance Feedback for Sales Lead Ranking

DGX agent

arXiv:2607.20655v1 Announce Type: cross Abstract: Lead ranking in Customer Relationship Management (CRM) systems faces a persistent challenge: models achieving high offline accuracy often underperform

safetyarxiv-cs-ai
24 Jul 2026
Model Releases

SciExplore: Evaluating Autonomous Agents from Scientific Navigation to Information Integration

DGX agent

arXiv:2607.20926v1 Announce Type: new Abstract: Scientific research involves complex information-seeking and reasoning workflows across heterogeneous sources. However, existing benchmarks primarily em

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

SkillCorpus: Consolidating and Evaluating the Open Skill Ecosystem for Real-World LLM Agents

DGX agent

arXiv:2607.15557v4 Announce Type: replace Abstract: Agent skills, SKILL files that package reusable procedural knowledge for an LLM agent, are a popular mechanism for extending agent capabilities. Pub

model-releasesarxiv-cs-cl
24 Jul 2026
Model Releases

SOAP, Muon, and Beyond: Pushing LLM Pretraining Scales

DGX agent

arXiv:2607.20548v1 Announce Type: cross Abstract: Higher-order optimizers such as Muon and SOAP offer faster convergence than AdamW, but their computational cost and numerical stability challenges hav

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

SonicSampler: Unified Tile-Aware Kernels for LLM Sampling and Speculative Verification

DGX agent

arXiv:2607.20475v1 Announce Type: new Abstract: Sampling in LLM inference comprises a combinatorial set of logit processing, token selection, and verification operations for speculative decoding. Howe

model-releasesarxiv-cs-ai
24 Jul 2026
Research

Sparse Concept Channels in Frozen 3D CT Vision Encoders

DGX agent

arXiv:2607.20993v1 Announce Type: cross Abstract: Large vision-language models are becoming increasingly dominant in 3D medical image interpretation, but we rarely know which internal units encode cli

researcharxiv-cs-ai
24 Jul 2026
Research

Spectral Concentration and Recovery in Sparse High-Dimensional Random Geometric Graphs

DGX agent

arXiv:2607.14304v2 Announce Type: replace-cross Abstract: We study sparse threshold random geometric graphs generated by high-dimensional spherical or Gaussian latent vectors. Although each edge has m

researcharxiv-cs-lg
24 Jul 2026
Tutorials

SR-TTT Does Not Learn Retrieval: A Correction and Mechanistic Post-Mortem of Surprisal-Aware Residual Test-Time Training

DGX agent

arXiv:2603.06642v2 Announce Type: replace-cross Abstract: Test-Time Training (TTT) language models replace the KV-cache with fast weights updated during inference, achieving O(1) memory but suffering

tutorialsarxiv-cs-ai
24 Jul 2026
Research

Tackling Heterogeneity in Federated Learning via Variance-Reduced Boltzmann Sampling within Homogeneous Social Coalitions

DGX agent

arXiv:2506.02897v3 Announce Type: replace Abstract: Federated Learning (FL) enables privacy-preserving collaborative model training, but its effectiveness is often limited by client data heterogeneity

researcharxiv-cs-lg
24 Jul 2026
Model Releases

Telco-GAIA: Bilingual Benchmark for Agents in Telecom Domain

DGX agent

arXiv:2607.20510v1 Announce Type: new Abstract: We introduce Telco-GAIA, a bilingual, multi-modal benchmark for evaluating tool-using agents on the data of a real-world telecommunications operator. Te

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

The Hidden Footprint: Making Storage a First-Class Metric for LLM Agent Evaluation

DGX agent

arXiv:2607.11149v3 Announce Type: replace Abstract: LLM agent benchmarks measure task completion, reliability, and inference cost, but not the persistent data an agent run leaves on disk, including lo

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

The Second LoViF 2026 Challenge on Real-World All-in-One Image Restoration: Methods and Results

DGX agent

arXiv:2607.21118v1 Announce Type: new Abstract: This paper presents a review of the second LoViF Challenge on Real-World All-in-One Image Restoration. The challenge aims to advance unified image resto

model-releasesarxiv-cs-cv
24 Jul 2026
Model Releases

TOUR: A Trajectory-Level Unlearning Benchmark for Offline Reinforcement Learning

DGX agent

arXiv:2607.21111v1 Announce Type: cross Abstract: Offline Reinforcement Learning (RL) agents are trained on fixed behavioral trajectories, which makes trajectory-level deletion important when selected

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

TransBiolab: A Real-World Multi-View Dataset of Cluttered Transparent Biomedical Objects

DGX agent

arXiv:2607.21071v1 Announce Type: new Abstract: Autonomous biomedical laboratories increasingly rely on visual perception to recognize, localize, and manipulate transparent plasticware, yet high-quali

model-releasesarxiv-cs-cv
24 Jul 2026
Model Releases

U-CFR: Uncertainty-Guided Cascade Forward Refinement for Interactive Segmentation

DGX agent

arXiv:2607.20705v1 Announce Type: cross Abstract: Interactive image segmentation is critical for efficient image annotation; however, existing methods often require many corrective clicks or rely on p

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Visual Contrastive Self-Distillation

DGX agent

arXiv:2607.21556v1 Announce Type: cross Abstract: On-policy self-distillation (OPSD) is promising as it removes the external teacher required by on-policy distillation (OPD), yet it still needs asymme

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

VPWEM: Non-Markovian Visuomotor Policy with Working and Episodic Memory

DGX agent

arXiv:2603.04910v2 Announce Type: replace-cross Abstract: Imitation learning from human demonstrations has achieved significant success in robotic control, yet most visuomotor policies still condition

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

When Does Recurrence Become an Algorithm? Convergence Selection in Weight-Tied Looped Transformers

DGX agent

arXiv:2607.20594v1 Announce Type: cross Abstract: When does a weight-tied looped transformer -- one block applied T times -- implement an actual algorithm? We answer with four findings from controlled

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning

DGX agent

arXiv:2607.09328v2 Announce Type: replace-cross Abstract: Answering complex questions over long documents frequently requires integrating evidence that the source itself disperses naturally across dis

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

ZONDA: Zero-shot Object Navigation with Dynamic Avoidance in Multi-floor Environments

DGX agent

arXiv:2607.21025v1 Announce Type: new Abstract: In Object Goal Navigation task, existing methods are typically restricted to static and single-floor environments, ignoring cross-floor topologies and d

model-releasesarxiv-cs-ro
24 Jul 2026
Research

Attention Without Grounding: Causal Evaluation of Visual Explanations in Medical VLMs

DGX agent

arXiv:2607.18577v1 Announce Type: new Abstract: Attention and saliency heatmaps are widely used to explain medical Vision-Language Model (VLM) outputs on chest X-rays, yet whether they truly highlight

researcharxiv-cs-cv
23 Jul 2026
Model Releases

Boundary Embedding Shaping with Adaptive Contrastive Learning for Graph Structural Disentanglement

DGX agent

arXiv:2606.20283v2 Announce Type: replace-cross Abstract: Graph neural networks (GNNs) excel at aggregating neighbor information for classification, yet their performance is hindered by graph structur

model-releasesarxiv-cs-ai
23 Jul 2026
Applications

Challenges of Explainability in Continual Learning for Time Series Forecasting

DGX agent

arXiv:2607.19382v1 Announce Type: cross Abstract: Deep learning models have shown strong potential for time series forecasting, yet their deployment in real-world environmental monitoring remains chal

applicationsarxiv-cs-ai
23 Jul 2026
Research

Classical Hardware Acceleration of Quantum Autoencoders for Real-Time Anomaly Detection in Collider Experiments

DGX agent

arXiv:2607.20302v1 Announce Type: new Abstract: Quantum machine learning (QML) algorithms in high energy physics (HEP) can efficiently represent and leverage long-range, high-order correlations in hig

researcharxiv-cs-lg
23 Jul 2026
Model Releases

Closing the Lab-to-Store Gap: A Data-Efficient Post-Training and Experience-Driven Learning VLA Framework for Retail Humanoids

DGX agent

arXiv:2607.20345v1 Announce Type: cross Abstract: Closing the gap between benchmark performance and reliable real-world operation remains a central challenge for Vision-Language-Action (VLA) humanoid

model-releasesarxiv-cs-ai
23 Jul 2026
Safety

Co-Evolving LLM Evaluators and Policies via DynamicRubric

DGX agent

arXiv:2607.20083v1 Announce Type: cross Abstract: Post-training with evaluator feedback on policy-induced samples serves as a major mechanism for improving large language models. As policies improve,

safetyarxiv-cs-ai
23 Jul 2026
Model Releases

contrib: allow all AI-generated code in general by ngxson · Pull Request #26012 · ggml-org/llama.cpp

DGX agent

Having read some merged PRs in the past, I know that they were fully written by Claude Code (or similar), so this basically fixes the delusion. But at the same time, we might start seeing more AI slop

model-releasesr-localllama
23 Jul 2026
Model Releases

CR-Refiner: An Object-Centric Optimal Transport Reranker for Edit-Conditioned 3D Scene Retrieval

DGX agent

arXiv:2607.19115v1 Announce Type: new Abstract: Edit-conditioned 3D scene retrieval pairs a reference 3D room with a natural-language modification and retrieves rooms from a corpus that satisfy the ed

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

Diffusion ReRoll: Revisable Denoising for Robotic Sequential Prediction

DGX agent

arXiv:2607.19919v1 Announce Type: cross Abstract: We propose Diffusion ReRoll, a diffusion-based framework for robotic sequential prediction that enables revisable denoising over horizons. Existing di

model-releasesarxiv-cs-lg
23 Jul 2026
Model Releases

DocOps: A Verifiable Benchmark for Autonomous Agents in Complex Document Operations

DGX agent

arXiv:2607.19865v1 Announce Type: new Abstract: As autonomous agents rapidly evolve, their ability to reliably manipulate ubiquitous digital documents has become critical for enabling general-purpose

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

DocShield: Towards AI Document Safety via Evidence-Grounded Agentic Reasoning

DGX agent

arXiv:2604.02694v2 Announce Type: replace-cross Abstract: The rapid progress of generative AI has enabled increasingly realistic text-centric image forgeries, posing major challenges to document safet

model-releasesarxiv-cs-ai
23 Jul 2026
Research

Domain-Adapted Power Curve for Cross-Farm Applications

DGX agent

arXiv:2607.19744v1 Announce Type: cross Abstract: The wind energy industry relies on accurate power curve models to make power forecast, evaluate turbine performance, quantify upgrade, or support site

researcharxiv-cs-lg
23 Jul 2026
Model Releases

DQAOA-GPT: AI-Accelerated Distributed Quantum Optimization for Combinatorial Problems

DGX agent

arXiv:2607.20225v1 Announce Type: cross Abstract: While combinatorial optimization problems are central to many scientific and engineering applications, their solution remains challenging due to expon

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

DREMnet: An Interpretable Denoising Framework for Semi-Airborne Transient Electromagnetic Signal

DGX agent

arXiv:2503.22223v2 Announce Type: replace Abstract: The semi-airborne transient electromagnetic method (SATEM) is capable of conducting rapid surveys over large-scale and hard-to-reach areas. However,

model-releasesarxiv-cs-lg
23 Jul 2026
Model Releases

Emergent Autonomous Drifting for Collision Avoidance in Real-World Winter Driving Scenarios

DGX agent

arXiv:2607.19484v1 Announce Type: new Abstract: Real-world collision avoidance is a core motivation for studying the dynamics and control of high sideslip drifting in vehicles, yet the practical benef

model-releasesarxiv-cs-ro
23 Jul 2026
Agents

Environment-free Synthetic Data Generation for API-Calling Agents

DGX agent

arXiv:2607.16900v2 Announce Type: replace Abstract: Training API-calling large language model (LLM) agents demands massive amounts of high-quality trajectories. However, collecting such data at scale

agentsarxiv-cs-ai
23 Jul 2026
Model Releases

ETPDesigner: Multi-Agent Orchestration for Interactive Multimodal Electronic Theater Program

DGX agent

arXiv:2607.19947v1 Announce Type: new Abstract: Electronic Theater Programs (ETPs) serve as critical promotional media in the performing arts, comprising a multi-page collection of heterogeneous visua

model-releasesarxiv-cs-cv
23 Jul 2026
Local Ai

From Classification to Localization and Clinical Validation: Large-Scale Development of a Deep Learning System for Thoracic Disease Detection on Chest Radiographs in Thailand

DGX agent

arXiv:2607.09305v2 Announce Type: replace-cross Abstract: Chest radiography (CXR) remains the most widely used thoracic imaging modality, yet expert interpretation is constrained by a severe shortage

local-aiarxiv-cs-lg
23 Jul 2026
Research

HijackKV: New Threat in Position-Independent KV Cache Reuse

DGX agent

arXiv:2607.19957v1 Announce Type: cross Abstract: Key-Value (KV) cache reduces inference latency in large language models (LLMs). Traditional prefix-based reuse has low cache hit rates across inferenc

researcharxiv-cs-ai
23 Jul 2026
Industry

Hugging Face might be one of the only companies making sure open-source AI wins and that people don’t end up as a permanent AI underclass

DGX agent

Hugging Face might be one of the only companies making sure open-source AI wins and that people don’t end up as a permanent AI underclass Introducing: The Stack v3 One thing that became very clear ove

industryclem-delangue--x
23 Jul 2026
Model Releases

Hypothesis-and-Refinement Learning of Organic Structures from Multimodal Spectroscopic Data

DGX agent

arXiv:2607.19816v1 Announce Type: cross Abstract: Determining molecular structures from spectroscopic data remains fundamentally challenging because the inverse problem is intrinsically underdetermine

model-releasesarxiv-cs-lg
23 Jul 2026
Local Ai

In-the-Flow Agentic System Optimization for Effective Planning and Tool Use

DGX agent

arXiv:2510.05592v2 Announce Type: replace Abstract: Outcome-driven reinforcement learning has advanced reasoning in large language models (LLMs), but prevailing tool-augmented approaches train a singl

local-aiarxiv-cs-ai
23 Jul 2026
Research

InstructMixup: Instruction-Guided Salient Patch Editing for Robust Data Augmentation

DGX agent

arXiv:2607.19324v1 Announce Type: new Abstract: In image and video technologies, data augmentation is widely used to improve the generalization of deep visual models, and mixup-based strategies that i

researcharxiv-cs-cv
23 Jul 2026
Safety

Interpretable Fuzzy Rule-Based Regression Extension for Ex-Fuzzy Library

DGX agent

arXiv:2607.20277v1 Announce Type: new Abstract: Machine learning models achieve high predictive accuracy in regression tasks, but their deployment in safety-critical and regulated domains requires int

safetyarxiv-cs-lg
23 Jul 2026
Research

Kontinuous Kontext: Continuous Strength Control for Instruction-based Image Editing

DGX agent

arXiv:2510.08532v2 Announce Type: replace Abstract: Instruction-based image editing offers a powerful and intuitive way to manipulate images through natural language. Yet, relying solely on text instr

researcharxiv-cs-cv
23 Jul 2026
Research

Learning About Learning: A Path from Spin Glasses to Artificial Intelligence

DGX agent

arXiv:2601.07635v3 Announce Type: replace-cross Abstract: The Hopfield model, originally inspired by spin glasses, occupies a central place at the intersection of statistical mechanics, neural network

researcharxiv-cs-ai
23 Jul 2026
← Previous
1…670671672673674…1369
Next →