AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
AllBlog
87,617Total entries
1Added by human
87,616Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,987 results
Research

Optimal Query Allocation in Extractive QA with LLMs: A Learning-to-Defer Framework with Theoretical Guarantees

DGX agent

arXiv:2410.15761v4 Announce Type: replace Abstract: Large Language Models excel in generative tasks but exhibit inefficiencies in structured text selection, particularly in extractive question answeri

researcharxiv-cs-cl
21 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

Playing Devil's Advocate: Off-the-Shelf Persona Vectors Rival Targeted Steering for Sycophancy

DGX agent

arXiv:2605.21006v1 Announce Type: cross Abstract: We study the effect of different persona on extbf{sycophancy}: model's agreement with users even when the user is incorrect. The standard mitigation,

researcharxiv-cs-cl
21 May 2026
Model Releases

Point Cloud Sequence Encoding for Material-conditioned Graph Network Simulators

DGX agent

arXiv:2605.20978v1 Announce Type: new Abstract: Graph Network Simulators (GNSs) have emerged as powerful surrogates for complex physics-based simulation, offering inherent differentiability and orders

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Pseudo-Formalization for Automatic Proof Verification

DGX agent

arXiv:2605.20531v1 Announce Type: cross Abstract: Reliable verification of proofs remains a bottleneck for training and evaluating AI systems on hard mathematical reasoning. Fully formal proofs, in la

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Query-Calibrated Segmental Admission for Descriptor-Agnostic LiDAR Loop Closure in Repetitive Environments

DGX agent

arXiv:2512.09447v2 Announce Type: replace-cross Abstract: Structurally repetitive environments produce visually plausible but aliased LiDAR loop candidates that can destabilize pose-graph optimization

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

RCGDet3D: Rethinking 4D Radar-Camera Fusion-based 3D Object Detection with Enhanced Radar Feature Encoding

DGX agent

arXiv:2605.21112v1 Announce Type: new Abstract: 4D automotive radar is indispensable for autonomous driving due to its low cost and robustness, yet its point cloud sparsity challenges 3D object detect

model-releasesarxiv-cs-cv
21 May 2026
Safety

Regulating Anatomy-Aware Rewards via Trajectory-Integral Feedback for Volumetric Computed Tomography Analysis

DGX agent

arXiv:2605.20277v1 Announce Type: new Abstract: Medical vision-language models (VLMs) have rapidly advanced as general-purpose multimodal assistants, yet their deployment in 3D Computed Tomography (CT

safetyarxiv-cs-cv
21 May 2026
Model Releases

Residual Paving: Diagnosing the Routing Bottleneck in Selective Refusal Editing

DGX agent

arXiv:2605.20262v1 Announce Type: new Abstract: We study selective refusal editing as a three-way control problem: induce non-refusal on designated edit prompts while preserving benign behavior and ha

model-releasesarxiv-cs-lg
21 May 2026
Agents

Retrieval-Augmented Code Generation: A Survey with Focus on Repository-Level Approaches

DGX agent

arXiv:2510.04905v3 Announce Type: replace-cross Abstract: Recent advances in large language models (LLMs) have significantly improved automated code generation. While existing approaches have achieved

agentsarxiv-cs-cl
21 May 2026
Model Releases

Retrieval-Augmented Long-Context Translation for Cultural Image Captioning: Gators submission for AmericasNLP 2026 shared task

DGX agent

arXiv:2605.20626v1 Announce Type: new Abstract: We present the University of Florida Gators submission to the AmericasNLP 2026 shared task on cultural image captioning for Indigenous languages. Our tw

model-releasesarxiv-cs-cl
21 May 2026
Research

ReversedQ: Opportunities for Faster Q-Learning in Episodic Online Reinforcement Learning

DGX agent

arXiv:2605.20592v1 Announce Type: new Abstract: We study model-free Q-learning in finite-horizon episodic Markov Decision Processes (MDPs) with stationary dynamics across episodes. We identify a centr

researcharxiv-cs-lg
21 May 2026
Model Releases

roto 2.0: The Robot Tactile Olympiad

DGX agent

arXiv:2605.21429v1 Announce Type: cross Abstract: Tactile-based reinforcement learning (RL) is currently hindered by fragmented research and a focus on over-saturated orientation tasks. We introduce v

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Safety-Critical Control for Smoothed Implicit Contact Dynamics

DGX agent

arXiv:2605.21138v1 Announce Type: new Abstract: Smoothed implicit contact dynamics enables gradient-based planning and control for contact-rich tasks without predefined mode sequences. However, safety

model-releasesarxiv-cs-ro
21 May 2026
Model Releases

Seems GPT-5.2 reaches expert level in peer review: 45 scientists took 469 hours evaluating human & AI reviews on 82 papers. 'Surprisingly, c…

DGX agent

Seems GPT-5.2 reaches expert level in peer review: 45 scientists took 469 hours evaluating human & AI reviews on 82 papers. 'Surprisingly, current AI reviewers are competitive even with the top-rated

model-releasesethan-mollick--x
21 May 2026
Model Releases

Semiparametric Efficient Bilevel Gradient Estimation

DGX agent

arXiv:2605.21341v1 Announce Type: cross Abstract: Functional bilevel methods estimate a lower-level function and plug it into a hypergradient, but this plug-in gradient can retain first-order bias whe

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

ShapeBench: A Scalable Benchmark and Diagnostic Suite for Standardized Evaluation in Aerodynamic Shape Optimization

DGX agent

arXiv:2605.20763v1 Announce Type: new Abstract: Rapid progress in aerodynamic shape optimization (ASO) has outpaced currently-available standardized evaluation frameworks. Fair comparison requires a u

model-releasesarxiv-cs-lg
21 May 2026
Safety

Shiny Stories, Hidden Struggles: Investigating the Representation of Disability Through the Lens of LLMs

DGX agent

arXiv:2605.20191v1 Announce Type: new Abstract: Modern Large Language Models (LLMs) have recently attracted much attention for their ability to simulate human behavior and generate text that reflects

safetyarxiv-cs-cl
21 May 2026
Research

SmoCap: Unified Scale-Pose Canonicalization with Proxy-Mapped Trust-Region QP

DGX agent

arXiv:2605.20850v1 Announce Type: new Abstract: Objective: Stage-wise workflows that separate model scaling and inverse kinematics can induce morphology-posture compensation, resulting in anatomically

researcharxiv-cs-ro
21 May 2026
Model Releases

SOPE: Stabilizing Off-Policy Evaluation for Online RL with Prior Data

DGX agent

arXiv:2605.05863v2 Announce Type: replace Abstract: Incorporating prior data into online reinforcement learning accelerates training but typically forces a difficult trade-off between high computation

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Spotify Labs launches Studio, a NotebookLM-like desktop app to generate private, AI-powered podcasts, in research preview across 20+ markets (Ivan Mehta/TechCrunch)

DGX agent

Ivan Mehta / TechCrunch: Spotify Labs launches Studio, a NotebookLM-like desktop app to generate private, AI-powered podcasts, in research preview across 20+ markets — One of the common features for c

model-releasestechmeme
21 May 2026
Model Releases

Spotify says it has 1M+ subscriptions to Audiobook+, which is on track for $100M in ARR, and unveils an ElevenLabs-powered audiobook self-publishing tool (Ivan Mehta/TechCrunch)

DGX agent

Ivan Mehta / TechCrunch: Spotify says it has 1M+ subscriptions to Audiobook+, which is on track for $100M in ARR, and unveils an ElevenLabs-powered audiobook self-publishing tool — Alongside tools for

model-releasestechmeme
21 May 2026
Safety

Statistical Guarantees in the Search for Less Discriminatory Algorithms

DGX agent

arXiv:2512.23943v2 Announce Type: replace-cross Abstract: U.S. discrimination law can impose liability on firms that fail to adopt a less discriminatory alternative (LDA): a decision policy that achie

safetyarxiv-cs-lg
21 May 2026
Safety

STiTch: Semantic Transition and Transportation in Collaboration for Training-Free Zero-Shot Composed Image Retrieval

DGX agent

arXiv:2605.21261v1 Announce Type: new Abstract: Training-free zero-shot composed image retrieval models are recently gaining increasing research interest due to their generalizability and flexibility

safetyarxiv-cs-cv
21 May 2026
Research

TabPFN-MT: A Natively Multitask In-Context Learner for Tabular Data

DGX agent

arXiv:2605.20234v1 Announce Type: new Abstract: Prior-Data Fitted networks (PFNs) have been very successful in tabular contexts, handling prediction tasks in context. However, they are designed for si

researcharxiv-cs-lg
21 May 2026
Model Releases

Theoretical guidelines for annealed Langevin dynamics in compositional simulation-based inference

DGX agent

arXiv:2605.21253v1 Announce Type: cross Abstract: Compositional score-based approaches to simulation-based inference (SBI) approximate the posterior over a shared parameter given n independent observa

model-releasesarxiv-cs-lg
21 May 2026
Safety

Time-Prompt: Integrated Heterogeneous Prompts for Unlocking LLMs in Time Series Forecasting

DGX agent

arXiv:2506.17631v4 Announce Type: replace Abstract: Time series forecasting aims to model temporal dependencies among variables for future state inference, holding significant importance and widesprea

safetyarxiv-cs-lg
21 May 2026
Research

Tiny-Engram: Trigger-Indexed Concept Tables for Generative Vision

DGX agent

arXiv:2605.20309v1 Announce Type: new Abstract: Current personalization methods for generative vision models typically encode new concepts through continuous adapters or weight updates, yet provide li

researcharxiv-cs-cv
21 May 2026
Model Releases

Token Maxxing, April 2026-May 2026, RIP 🪦

DGX agent

Token Maxxing, April 2026-May 2026, RIP 🪦 🦔Microsoft canceled its internal Claude Code licenses this week after token-based billing made the cost untenable, even for a company with effectively infinit

model-releasesgary-marcus--x
21 May 2026
Agents

Towards Resilient and Autonomous Networks: A BlueSky Vision on AI-Native 6G

DGX agent

arXiv:2605.21395v1 Announce Type: cross Abstract: The proliferation of emerging applications, such as autonomous driving and immersive experiences, demands cellular networks that are not only faster,

agentsarxiv-cs-lg
21 May 2026
Model Releases

Towards UAV Detection in the Real World: A New Multispectral Dataset UAVNet-MS and a New Method

DGX agent

arXiv:2605.20963v1 Announce Type: new Abstract: The proliferation of unmanned aerial vehicles (UAVs) has created urgent demand for precise UAV monitoring. Existing RGB-based systems rely on spatial cu

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Training distribution determines the ceiling of drug-blind cancer sensitivity prediction

DGX agent

arXiv:2605.20885v1 Announce Type: new Abstract: Precision oncology requires predicting which drugs will suppress a specific tumor from its molecular profile, but drug-blind sensitivity prediction has

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

TRAM: Test-Time Risk Adaptation with Mixture of Agents

DGX agent

arXiv:2408.08812v2 Announce Type: replace Abstract: Deployed reinforcement learning agents often face safety requirements that are specified only after training, such as new hazard maps, revised risk

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Try Composer 2.5

DGX agent

Try Composer 2.5 New CursorBench results just dropped. Two big takeaways. Composer 2.5 is way better than most people think. 63.2% score at $0.55 per task. Nearly matching Opus 4.7 Max and GPT 5.5 Ext

model-releaseselon-musk--x
21 May 2026
Hardware

Understanding Deterioration Random Effects for Causal Discovery in Infrastructure Management

DGX agent

arXiv:2605.20400v1 Announce Type: cross Abstract: Infrastructure deterioration poses significant challenges for asset management, yet existing approaches rely on population-averaged models that overlo

hardwarearxiv-cs-lg
21 May 2026
Research

VIHD: Visual Intervention-based Hallucination Detection for Medical Visual Question Answering

DGX agent

arXiv:2605.20772v1 Announce Type: new Abstract: While medical Multimodal Large Language Models (MLLMs) have shown promise in assisting diagnosis, they still frequently generate hallucinated responses

researcharxiv-cs-cv
21 May 2026
Model Releases

VISTA: Technical Report for the Ego4D Short-Term Object Interaction Anticipation at EgoVis 2026

DGX agent

arXiv:2605.20901v1 Announce Type: new Abstract: We propose VISTA, a V-JEPA Integrated StillFast Temporal Anticipator for the Ego4D Short-Term Object Interaction Anticipation (STA) Challenge at EgoVis

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

What Do Biomedical NER and Entity Linking Benchmarks Measure? A Corpus-Centric Diagnostic Framework

DGX agent

arXiv:2605.20537v1 Announce Type: new Abstract: Biomedical named entity recognition (NER) and entity linking (EL) strongly depend on annotated corpora, but the utility of these resources for benchmark

model-releasesarxiv-cs-cl
21 May 2026
Agents

Why Latent Actions Fail, and How to Prevent It

DGX agent

arXiv:2605.20223v1 Announce Type: new Abstract: Latent action models (LAMs) aim to learn action-like representations from unlabeled videos by compressing frame-to-frame changes. The frames of in-the-w

agentsarxiv-cs-cv
21 May 2026
Model Releases

Wordle 1,796 4/6 ⬛⬛⬛🟨🟨 ⬛⬛⬛⬛⬛ ⬛⬛🟩⬛🟨 🟩🟩🟩🟩🟩

DGX agent

This post shows a Wordle game result where the player solved puzzle #1,796 in 4 attempts, with the final answer being a five-letter word with the pattern shown in green squares. The emoji grid display

model-releasesanthropic--x
21 May 2026
Model Releases

ZEBRA: Zero-shot Budgeted Resource Allocation for LLM Orchestration

DGX agent

arXiv:2605.20485v1 Announce Type: new Abstract: As autonomous agents increasingly execute end-to-end tasks under fixed monetary budgets, the pressing open question shifts from whether the budget is re

model-releasesarxiv-cs-lg
21 May 2026
Agents

1Password extends OpenAI collaboration with Codex MCP server for just-in-time credential access

DGX agent

Cybersecurity and password service provider 1Password LLC today expanded its collaboration with OpenAI Group PBC, releasing a Model Context Protocol server that lets the Codex coding agent pull creden

agentssiliconangle
20 May 2026
Model Releases

A Case for Agentic Tuning: From Documentation to Action in PostgreSQL

DGX agent

arXiv:2605.19988v1 Announce Type: cross Abstract: Documentation has long guided computer system tuning by distilling expert knowledge into per-parameter recommendations. Yet such guides capture only w

model-releasesarxiv-cs-ai
20 May 2026
Safety

A Geometric Analysis of Sign-Magnitude Asymmetry in a ReLU + RMSNorm Block under Ternary Quantization

DGX agent

arXiv:2605.18933v1 Announce Type: new Abstract: Pre-norm Transformers with RMSNorm tolerate ternary {-1,0,+1} weight quantization with surprisingly small loss (Ma et al., 2024). We give a geometric ex

safetyarxiv-cs-lg
20 May 2026
Local Ai

A glimpse into the fleet: Deep space patrol (4K Sci-Fi) [OC]

DGX agent

A user-generated AI image creation showcasing a deep space patrol scene rendered in 4K quality using Stable Diffusion, a popular text-to-image generation model. The post was shared on the r/StableDiff

local-air-stablediffusion
20 May 2026
Research

A Nash Equilibrium Framework For Training-Free Multimodal Step Verification

DGX agent

arXiv:2605.20033v1 Announce Type: new Abstract: Multimodal large language models often generate reasoning chains containing subtle errors that lead to incorrect answers. Current verification approache

researcharxiv-cs-cv
20 May 2026
Model Releases

A Nonlinear Complexity Index for Wearable PPG Cardiovascular Stability: Multiscale Validation, Systematic Evaluation Correction, and Bayesian Parameter Optimization

DGX agent

arXiv:2605.18802v1 Announce Type: cross Abstract: Cardiovascular stability estimation from wearable photoplethysmography (PPG) requires a principled nonlinear framework, yet major gaps persist in heur

model-releasesarxiv-cs-ai
20 May 2026
Hardware

Accelerating Sparse Transformer Inference on GPU

DGX agent

arXiv:2506.06095v4 Announce Type: replace Abstract: Large language models (LLMs) are popular around the world due to their powerful understanding capabilities. As the core component of LLMs, accelerat

hardwarearxiv-cs-lg
20 May 2026
Model Releases

Adapted Center and Scale Prediction: More Stable and More Accurate

DGX agent

arXiv:2002.09053v3 Announce Type: replace Abstract: Pedestrian detection benefits from deep learning technology and gains rapid development in recent years. Most of detectors follow general object det

model-releasesarxiv-cs-cv
20 May 2026
← Previous
1…848849850851852…1313
Next →