AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
Human
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
Model Releases

ProRank: Prompt Warmup via Reinforcement Learning for Small Language Models Reranking

DGX agent

arXiv:2506.03487v3 Announce Type: replace-cross Abstract: Reranking is fundamental to information retrieval and retrieval-augmented generation, with recent Large Language Models (LLMs) significantly a

model-releasesarxiv-cs-cl
17 Apr 2026
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

ProRe: A Proactive Reward System for GUI Agents via Reasoner-Actor Collaboration

DGX agent

arXiv:2509.21823v2 Announce Type: replace Abstract: Reward is critical to the evaluation and training of large language models (LLMs). However, existing rule-based or model-based reward methods strugg

safetyarxiv-cs-ai
17 Apr 2026
Safety

PROXIMA: A Reliability Scoring Framework for Proxy Metrics in Online Controlled Experiments

DGX agent

arXiv:2604.14352v1 Announce Type: cross Abstract: Online A/B testing at scale relies on proxy metrics -- short-term, easily-measured signals used in place of slow-moving long-term outcomes. When the p

safetyarxiv-cs-lg
17 Apr 2026
Tutorials

Pruning Long Chain-of-Thought of Large Reasoning Models via Small-Scale Preference Optimization

DGX agent

arXiv:2508.10164v2 Announce Type: replace Abstract: Recent advances in Large Reasoning Models (LRMs) have demonstrated strong performance on complex tasks through long Chain-of-Thought (CoT) reasoning

tutorialsarxiv-cs-ai
17 Apr 2026
Research

Psychological Steering of Large Language Models

DGX agent

arXiv:2604.14463v1 Announce Type: new Abstract: Large language models (LLMs) emulate a consistent human-like behavior that can be shaped through activation-level interventions. This paradigm is conver

researcharxiv-cs-cl
17 Apr 2026
Research

PUFFIN: Protein Unit Discovery with Functional Supervision

DGX agent

arXiv:2604.14796v1 Announce Type: cross Abstract: Proteins carry out biological functions through the coordinated action of groups of residues organized into structural arrangements. These arrangement

researcharxiv-cs-lg
17 Apr 2026
Model Releases

Purging the Gray Zone: Latent-Geometric Denoising for Precise Knowledge Boundary Awareness

DGX agent

arXiv:2604.14324v1 Announce Type: new Abstract: Large language models (LLMs) often exhibit hallucinations due to their inability to accurately perceive their own knowledge boundaries. Existing abstent

model-releasesarxiv-cs-cl
17 Apr 2026
Safety

Pushing the Boundaries of Multiple Choice Evaluation to One Hundred Options

DGX agent

arXiv:2604.14634v1 Announce Type: new Abstract: Multiple choice evaluation is widely used for benchmarking large language models, yet near ceiling accuracy in low option settings can be sustained by s

safetyarxiv-cs-cl
17 Apr 2026
Research

Q-MambaIR: Accurate Quantized Mamba for Efficient Image Restoration

DGX agent

arXiv:2503.21970v3 Announce Type: replace Abstract: State-Space Models (SSMs) have attracted considerable attention in Image Restoration (IR) due to their ability to scale linearly sequence length whi

researcharxiv-cs-cv
17 Apr 2026
Safety

QU-NLP at ArchEHR-QA 2026: Two-Stage QLoRA Fine-Tuning of Qwen3-4B for Patient-Oriented Clinical Question Answering and Evidence Sentence Alignment

DGX agent

arXiv:2604.14175v1 Announce Type: new Abstract: We present a unified system addressing both Subtask 3 (answer generation) and Subtask 4 (evidence sentence alignment) of the ArchEHR-QA Shared Task. For

safetyarxiv-cs-cl
17 Apr 2026
Research

QualiaNet: An Experience-Before-Inference Network

DGX agent

arXiv:2604.14193v1 Announce Type: new Abstract: Human 3D vision involves two distinct stages: an Experience Module, where stereo depth is extracted relative to fixation, and an Inference Module, where

researcharxiv-cs-cv
17 Apr 2026
Applications

Quality-Aware Calibration for AI-Generated Image Detection in the Wild

DGX agent

arXiv:2604.15027v1 Announce Type: new Abstract: Significant progress has been made in detecting synthetic images, however most existing approaches operate on a single image instance and overlook a key

applicationsarxiv-cs-cv
17 Apr 2026
Model Releases

QuantCode-Bench: A Benchmark for Evaluating the Ability of Large Language Models to Generate Executable Algorithmic Trading Strategies

DGX agent

arXiv:2604.15151v1 Announce Type: new Abstract: Large language models have demonstrated strong performance on general-purpose programming tasks, yet their ability to generate executable algorithmic tr

model-releasesarxiv-cs-cl
17 Apr 2026
Research

Quantitative Approximation Rates for Group Equivariant Learning

DGX agent

arXiv:2602.20370v2 Announce Type: replace Abstract: The universal approximation theorem establishes that neural networks can approximate any continuous function on a compact set. Later works in approx

researcharxiv-cs-lg
17 Apr 2026
Research

Quantization of Spiking Neural Networks Beyond Accuracy

DGX agent

arXiv:2604.14487v1 Announce Type: new Abstract: Quantization is a natural complement to the sparse, event-driven computation of Spiking Neural Networks, reducing memory bandwidth and arithmetic cost f

researcharxiv-cs-lg
17 Apr 2026
Research

Quantum-inspired tensor networks in machine learning models

DGX agent

arXiv:2604.14287v1 Announce Type: new Abstract: Tensor networks were developed in the context of many-body physics as compressed representations of multiparticle quantum states. These representations

researcharxiv-cs-lg
17 Apr 2026
Model Releases

Query pipeline optimization for cancer patient question answering systems

DGX agent

arXiv:2412.14751v2 Announce Type: replace Abstract: Retrieval-augmented generation (RAG) mitigates hallucination in Large Language Models (LLMs) by using query pipelines to retrieve relevant external

model-releasesarxiv-cs-cl
17 Apr 2026
Safety

R3D: Revisiting 3D Policy Learning

DGX agent

arXiv:2604.15281v1 Announce Type: new Abstract: 3D policy learning promises superior generalization and cross-embodiment transfer, but progress has been hindered by training instabilities and severe o

safetyarxiv-cs-cv
17 Apr 2026
Research

RACER: Retrieval-Augmented Contextual Rapid Speculative Decoding

DGX agent

arXiv:2604.14885v1 Announce Type: new Abstract: Autoregressive decoding in Large Language Models (LLMs) generates one token per step, causing high inference latency. Speculative decoding (SD) mitigate

researcharxiv-cs-cl
17 Apr 2026
Safety

RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework

DGX agent

arXiv:2604.15308v1 Announce Type: new Abstract: High-level autonomous driving requires motion planners capable of modeling multimodal future uncertainties while remaining robust in closed-loop interac

safetyarxiv-cs-cv
17 Apr 2026
Research

Random Matrix Theory for Deep Learning: Beyond Eigenvalues of Linear Models

DGX agent

arXiv:2506.13139v2 Announce Type: replace-cross Abstract: Modern Machine Learning (ML) and Deep Neural Networks (DNNs) often operate on high-dimensional data and rely on overparameterized models, wher

researcharxiv-cs-lg
17 Apr 2026
Safety

RaTA-Tool: Retrieval-based Tool Selection with Multimodal Large Language Models

DGX agent

arXiv:2604.14951v1 Announce Type: cross Abstract: Tool learning with foundation models aims to endow AI systems with the ability to invoke external resources -- such as APIs, computational utilities,

safetyarxiv-cs-cl
17 Apr 2026
Safety

Reasoning Dynamics and the Limits of Monitoring Modality Reliance in Vision-Language Models

DGX agent

arXiv:2604.14888v1 Announce Type: new Abstract: Recent advances in vision language models (VLMs) offer reasoning capabilities, yet how these unfold and integrate visual and textual information remains

safetyarxiv-cs-cl
17 Apr 2026
Research

ReasonScaffold: A Scaffolded Reasoning-based Annotation Protocol for Human-AI Co-Annotation

DGX agent

arXiv:2603.21094v3 Announce Type: replace Abstract: Human annotation is central to NLP evaluation, yet subjective tasks often exhibit substantial variability across annotators. While large language mo

researcharxiv-cs-cl
17 Apr 2026
Safety

RECOVER: Designing a Large Language Model-based Remote Patient Monitoring System for Postoperative Gastrointestinal Cancer Care

DGX agent

arXiv:2502.05740v2 Announce Type: replace-cross Abstract: Cancer surgery is a key treatment for gastrointestinal (GI) cancers, a group of cancers that account for more than 35% of cancer-related death

safetyarxiv-cs-ai
17 Apr 2026
Hardware

Reference-Free Sampling-Based Model Predictive Control

DGX agent

arXiv:2511.19204v3 Announce Type: replace Abstract: We present a sampling-based model predictive control (MPC) framework that enables emergent locomotion without relying on handcrafted gait patterns o

hardwarearxiv-cs-ro
17 Apr 2026
Model Releases

Regret Tail Characterization of Optimal Bandit Algorithms with Generic Rewards

DGX agent

arXiv:2604.14876v1 Announce Type: cross Abstract: We study the tail behavior of regret in stochastic multi-armed bandits for algorithms that are asymptotically optimal in expectation. While minimizing

model-releasesarxiv-cs-lg
17 Apr 2026
Safety

Reinforcement Learning via Value Gradient Flow

DGX agent

arXiv:2604.14265v1 Announce Type: new Abstract: We study behavior-regularized reinforcement learning (RL), where regularization toward a reference distribution (the dataset in offline RL or the base m

safetyarxiv-cs-lg
17 Apr 2026
Model Releases

RELOAD: A Robust and Efficient Learned Query Optimizer for Database Systems

DGX agent

arXiv:2604.14725v1 Announce Type: cross Abstract: Recent advances in query optimization have shifted from traditional rule-based and cost-based techniques towards machine learning-driven approaches. A

model-releasesarxiv-cs-lg
17 Apr 2026
Tutorials

ReSS: Learning Reasoning Models for Tabular Data Prediction via Symbolic Scaffold

DGX agent

arXiv:2604.13392v1 Announce Type: new Abstract: Tabular data remains prevalent in high-stakes domains such as healthcare and finance, where predictive models are expected to provide both high accuracy

tutorialsarxiv-cs-ai
17 Apr 2026
Local Ai

Rethinking AI Hardware: A Three-Layer Cognitive Architecture for Autonomous Agents

DGX agent

arXiv:2604.13757v1 Announce Type: new Abstract: The next generation of autonomous AI systems will be constrained not only by model capability, but by how intelligence is structured across heterogeneou

local-aiarxiv-cs-ai
17 Apr 2026
Research

Rethinking LLM-Driven Heuristic Design: Generating Efficient and Specialized Solvers via Dynamics-Aware Optimization

DGX agent

arXiv:2601.20868v2 Announce Type: replace Abstract: Large Language Models (LLMs) have advanced the field of Combinatorial Optimization through automated heuristic generation. Instead of relying on man

researcharxiv-cs-lg
17 Apr 2026
Model Releases

Rethinking Patient Education as Multi-turn Multi-modal Interaction

DGX agent

arXiv:2604.14656v1 Announce Type: cross Abstract: Most medical multimodal benchmarks focus on static tasks such as image question answering, report generation, and plain-language rewriting. Patient ed

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

Retrieve, Then Classify: Corpus-Grounded Automation of Clinical Value Set Authoring

DGX agent

arXiv:2604.14616v1 Announce Type: new Abstract: Clinical value set authoring -- the task of identifying all codes in a standardized vocabulary that define a clinical concept -- is a recurring bottlene

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

ReviewGrounder: Improving Review Substantiveness with Rubric-Guided, Tool-Integrated Agents

DGX agent

arXiv:2604.14261v1 Announce Type: new Abstract: The rapid rise in AI conference submissions has driven increasing exploration of large language models (LLMs) for peer review support. However, LLM-base

model-releasesarxiv-cs-cl
17 Apr 2026
Research

Revisiting Token Compression for Accelerating ViT-based Sparse Multi-View 3D Object Detectors

DGX agent

arXiv:2604.14563v1 Announce Type: new Abstract: Vision Transformer (ViT)-based sparse multi-view 3D object detectors have achieved remarkable accuracy but still suffer from high inference latency due

researcharxiv-cs-cv
17 Apr 2026
Safety

Reward-Aware Trajectory Shaping for Few-step Visual Generation

DGX agent

arXiv:2604.14910v1 Announce Type: new Abstract: Achieving high-fidelity generation in extremely few sampling steps has long been a central goal of generative modeling. Existing approaches largely rely

safetyarxiv-cs-cv
17 Apr 2026
Model Releases

Right at My Level: A Unified Multilingual Framework for Proficiency-Aware Text Simplification

DGX agent

arXiv:2604.05302v2 Announce Type: replace Abstract: Text simplification supports second language (L2) learning by providing comprehensible input, consistent with the Input Hypothesis. However, constru

model-releasesarxiv-cs-cl
17 Apr 2026
Safety

RL-STPA: Adapting System-Theoretic Hazard Analysis for Safety-Critical Reinforcement Learning

DGX agent

arXiv:2604.15201v1 Announce Type: new Abstract: As reinforcement learning (RL) deployments expand into safety-critical domains, existing evaluation methods fail to systematically identify hazards aris

safetyarxiv-cs-lg
17 Apr 2026
Research

Robustness of Vision Foundation Models to Common Perturbations

DGX agent

arXiv:2604.14973v1 Announce Type: cross Abstract: A vision foundation model outputs an embedding vector for an image, which can be affected by common editing operations (e.g., JPEG compression, bright

researcharxiv-cs-cv
17 Apr 2026
Safety

RoSLAC: Robust Simultaneous Localization and Calibration of Multiple Magnetometers

DGX agent

arXiv:2604.14353v1 Announce Type: new Abstract: Localization of autonomous mobile robots (AMRs) in enclosed or semi-enclosed environments such as offices, hotels, hospitals, indoor parking facilities,

safetyarxiv-cs-ro
17 Apr 2026
Applications

Route to Rome Attack: Directing LLM Routers to Expensive Models via Adversarial Suffix Optimization

DGX agent

arXiv:2604.15022v1 Announce Type: cross Abstract: Cost-aware routing dynamically dispatches user queries to models of varying capability to balance performance and inference cost. However, the routing

applicationsarxiv-cs-cl
17 Apr 2026
Research

S2AM3D: Scale-controllable Part Segmentation of 3D Point Cloud

DGX agent

arXiv:2512.00995v3 Announce Type: replace Abstract: Part-level point cloud segmentation has recently attracted significant attention in 3D computer vision. Nevertheless, existing research is constrain

researcharxiv-cs-cv
17 Apr 2026
Safety

Safe Reinforcement Learning using Action Projection: Safeguard the Policy or the Environment?

DGX agent

arXiv:2509.12833v2 Announce Type: replace Abstract: Projection-based safety filters, which modify unsafe actions by mapping them to the closest safe alternative, are widely used to enforce safety cons

safetyarxiv-cs-lg
17 Apr 2026
Model Releases

SafeHarness: Lifecycle-Integrated Security Architecture for LLM-based Agent Deployment

DGX agent

arXiv:2604.13630v1 Announce Type: cross Abstract: The performance of large language model (LLM) agents depends critically on the execution harness, the system layer that orchestrates tool use, context

model-releasesarxiv-cs-ai
17 Apr 2026
Model Releases

SAGE Celer 2.6 Technical Card

DGX agent

arXiv:2604.14168v1 Announce Type: new Abstract: We introduce SAGE Celer 2.6, the latest in our line of general-purpose Celer models from SAGEA. Celer 2.6 is available in 5B, 10B, and 27B parameter siz

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

SAGE: Sign-Adaptive Gradient for Memory-Efficient LLM Optimization

DGX agent

arXiv:2604.07663v2 Announce Type: replace Abstract: The AdamW optimizer, while standard for LLM pretraining, is a critical memory bottleneck, consuming optimizer states equivalent to twice the model's

model-releasesarxiv-cs-lg
17 Apr 2026
Research

Sandwich: Joint Configuration Search and Hot-Switching for Efficient CPU LLM Serving

DGX agent

arXiv:2507.18454v2 Announce Type: replace-cross Abstract: CPUs are critical for LLM serving due to their availability, cost efficiency, and edge applicability. However, efficient CPU serving is hinder

researcharxiv-cs-ai
17 Apr 2026
← Previous
1…11511152115311541155…1247
Next →