AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
84,606 results
21 Apr 2026

OK, here's a resolution - I managed to get it to think using these settings: 'thinking': { 'type': 'adaptive', 'display': 'summarized' }, 'o…

ToolsDGX agent

OK, here's a resolution - I managed to get it to think using these settings: 'thinking': { 'type': 'adaptive', 'display': 'summarized' }, 'output_config': { 'effort': 'max' } Without 'display': 'summa

Omni-Embed-Audio: Leveraging Multimodal LLMs for Robust Audio-Text Retrieval

ApplicationsDGX agent

arXiv:2604.18360v1 Announce Type: cross Abstract: Audio-text retrieval systems based on Contrastive Language-Audio Pretraining (CLAP) achieve strong performance on traditional benchmarks; however, the

OmniHuman: A Large-scale Dataset and Benchmark for Human-Centric Video Generation

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub

arXiv:2604.18326v1 Announce Type: new Abstract: Recent advancements in audio-video joint generation models have demonstrated impressive capabilities in content creation. However, generating high-fidel

OmniVLA-RL: A Vision-Language-Action Model with Spatial Understanding and Online RL

SafetyDGX agent

arXiv:2604.17706v1 Announce Type: new Abstract: Visual-Language-Action (VLA) models represent a paradigm shift in embodied AI, yet existing frameworks often struggle with imprecise spatial perception,

OmniZip: Audio-Guided Dynamic Token Compression for Fast Omnimodal Large Language Models

ResearchDGX agent

arXiv:2511.14582v2 Announce Type: replace Abstract: Omnimodal large language models (OmniLLMs) have attracted increasing research attention of late towards unified audio-video understanding. However,

On Different Notions of Redundancy in Conditional-Independence-Based Discovery of Graphical Models

ResearchDGX agent

arXiv:2502.08531v3 Announce Type: replace Abstract: Conditional-independence-based discovery uses statistical tests to identify a graphical model that represents the independence structure of variable

On Inverse Problems, Parameter Estimation, and Domain Generalization

Model ReleasesDGX agent

arXiv:2506.06024v2 Announce Type: replace-cross Abstract: Signal restoration and inverse problems are key elements in most real-world data science applications. In the past decades, with the emergence

On-Orbit Space AI: Federated, Multi-Agent, and Collaborative Algorithms for Satellite Constellations

SafetyDGX agent

arXiv:2604.16518v1 Announce Type: new Abstract: Satellite constellations are transforming space systems from isolated spacecraft into networked, software-defined platforms capable of on-orbit percepti

On Safety Risks in Experience-Driven Self-Evolving Agents

SafetyDGX agent

arXiv:2604.16968v1 Announce Type: new Abstract: Experience-driven self-evolution has emerged as a promising paradigm for improving the autonomy of large language model agents, yet its reliance on self

On the Convergence and Size Transferability of Continuous-depth Graph Neural Networks

SafetyDGX agent

arXiv:2510.03923v2 Announce Type: replace Abstract: Continuous-depth graph neural networks, also known as Graph Neural Differential Equations (GNDEs), combine the structural inductive bias of Graph Ne

On the Emergence of Syntax by Means of Local Interaction

Model ReleasesDGX agent

arXiv:2604.17857v1 Announce Type: new Abstract: Can syntactic processing emerge spontaneously from purely local interaction? We present a concrete instance on a minimal system: an 18,658-parameter two

On the Generalization Bounds of Symbolic Regression with Genetic Programming

Model ReleasesDGX agent

arXiv:2604.17402v1 Announce Type: new Abstract: Symbolic regression (SR) with genetic programming (GP) aims to discover interpretable mathematical expressions directly from data. Despite its strong em

On the Importance and Evaluation of Narrativity in Natural Language AI Explanations

Model ReleasesDGX agent

arXiv:2604.18311v1 Announce Type: new Abstract: Explainable AI (XAI) aims to make the behaviour of machine learning models interpretable, yet many explanation methods remain difficult to understand. T

On the Importance of Tactile Sensing for Imitation Learning: A Case Study on Robotic Match Lighting

SafetyDGX agent

arXiv:2504.13618v4 Announce Type: replace Abstract: The field of robotic manipulation has advanced significantly in recent years. At the sensing level, several novel tactile sensors have been develope

On the Interpolation Effect of Score Smoothing in Diffusion Models

ResearchDGX agent

arXiv:2502.19499v3 Announce Type: replace Abstract: Diffusion models have achieved remarkable progress in various domains with an intriguing ability to produce new data that do not exist in the traini

On The Mathematics of the Natural Physics of Optimization

ResearchDGX agent

arXiv:2604.17645v1 Announce Type: cross Abstract: A number of optimization algorithms have been inspired by the physics of Newtonian motion. Here, we ask the question: do algorithms themselves obey so

On the Predictive Power of Representation Dispersion in Language Models

Model ReleasesDGX agent

arXiv:2506.24106v2 Announce Type: replace Abstract: We show that a language model's ability to predict text is tightly linked to the breadth of its embedding space: models that spread their contextual

On the Robustness of LLM-Based Dense Retrievers: A Systematic Analysis of Generalizability and Stability

ResearchDGX agent

arXiv:2604.16576v1 Announce Type: cross Abstract: Decoder-only large language models (LLMs) are increasingly replacing BERT-style architectures as the backbone for dense retrieval, achieving substanti

On the Sample Complexity of Learning for Blind Inverse Problems

ResearchDGX agent

arXiv:2512.23405v4 Announce Type: replace Abstract: Blind inverse problems arise in many experimental settings where both the signal of interest and the forward operator are (partially) unknown. In th

On the Shelf Life of Fine-Tuned LLM-Judges: Future-Proofing, Backward-Compatibility, and Question Generalization

SafetyDGX agent

arXiv:2509.23542v2 Announce Type: replace Abstract: The LLM-as-a-judge paradigm is widely used in both evaluating free-text model responses and reward modeling for model alignment and fine-tuning. Rec

On the Theory of Continual Learning with Gradient Descent for Neural Networks

ResearchDGX agent

arXiv:2510.05573v2 Announce Type: replace-cross Abstract: Continual learning, the ability of a model to adapt to an ongoing sequence of tasks without forgetting earlier ones, is a central goal of arti

One Adapts to Any: Meta Reward Modeling for Personalized LLM Alignment

SafetyDGX agent

arXiv:2601.18731v2 Announce Type: replace Abstract: Alignment of Large Language Models (LLMs) aims to align outputs with human preferences, and personalized alignment further adapts models to individu

one of my favorite quality of life features

Model ReleasesDGX agent

one of my favorite quality of life features Claude Code in the terminal will now show recaps when you switch focus away from the session and then come back. This should help you stay more in flow whil

One of the most jarring things about current AI is its lack of introspection ability and metacognition. It doesn't know what it doesn't know…

ResearchDGX agent

One of the most jarring things about current AI is its lack of introspection ability and metacognition. It doesn't know what it doesn't know, how it knows, or how it could find out. It's a one-way sys

One-Step Diffusion with Inverse Residual Fields for Unsupervised Industrial Anomaly Detection

ResearchDGX agent

arXiv:2604.18393v1 Announce Type: new Abstract: Diffusion models have achieved outstanding performance in unsupervised industrial anomaly detection (uIAD) by learning a manifold of normal data under t

OneDrive: Unified Multi-Paradigm Driving with Vision-Language-Action Models

AgentsDGX agent

arXiv:2604.17915v1 Announce Type: new Abstract: Vision-Language Models(VLMs) excel at autoregressive text generation, yet end-to-end autonomous driving requires multi-task learning with structured out

OneVL: One-Step Latent Reasoning and Planning with Vision-Language Explanation

AgentsDGX agent

arXiv:2604.18486v1 Announce Type: cross Abstract: Chain-of-Thought (CoT) reasoning has become a powerful driver of trajectory prediction in VLA-based autonomous driving, yet its autoregressive nature

Online Conformal Prediction with Adversarial Semi-bandit Feedback via Regret Minimization

SafetyDGX agent

arXiv:2604.17984v1 Announce Type: new Abstract: Uncertainty quantification is crucial in safety-critical systems, where decisions must be made under uncertainty. In particular, we consider the problem

Only the fraud was stopped by DOGE and, even then, only some of the fraud

Model ReleasesDGX agent

Only the fraud was stopped by DOGE and, even then, only some of the fraud On Friday, @StateDeptGHSD released its first wave of PEPFAR data, covering July-Sept 2025, after State took over USAID’s lifes

ONTO: A Token-Efficient Columnar Notation for LLM Input Optimization

ResearchDGX agent

arXiv:2604.17512v1 Announce Type: new Abstract: Serialization formats designed for document interchange impose structural overhead that becomes prohibitive when large language models consume operation

open-source community will work their magic!

IndustryDGX agent

open-source community will work their magic! If any inference providers could serve kimi k2.6 at 300+ tps reliably, I would be happy to switch from gpt 5.4 xhigh to kimi for daily tasks. It's a really

Open-TQ-Metal: Fused Compressed-Domain Attention for Long-Context LLM Inference on Apple Silicon

Model ReleasesDGX agent

arXiv:2604.16957v1 Announce Type: new Abstract: We present Open-TQ-Metal, the first implementation of fused compressed-domain attention on Apple Silicon, enabling 128K-context inference for Llama 3.1

OpenAI enables cost-per-click ads inside ChatGPT, setting bids at between 3 and 5 per click, in addition to CPMs (Digiday)

TutorialsDGX agent

Digiday: OpenAI enables cost-per-click ads inside ChatGPT, setting bids at between 3 and 5 per click, in addition to CPMs — The platform that can figure out how to prove AI chatbots can drive an outco

OpenAI launches ChatGPT Images 2.0, Codex Labs developer training service

Model ReleasesDGX agent

OpenAI Group PBC today launched ChatGPT Images 2.0, an upgraded version of the image generator built into its popular chatbot. The company also debuted a new technical training service called Codex La

OpenAI's new Euphony tool works almost exactly the same way as my Codex transcript viewer https://tools.simonwillison.net/codex-timeline?url…

ToolsDGX agent

OpenAI's new Euphony tool works almost exactly the same way as my Codex transcript viewer https://tools.simonwillison.net/codex-timeline?url=https%3A%2F%2Fgist.githubusercontent.com%2Fsimonw%2Fa9eb599

OpenAI’s updated image generator can now pull information from the web

IndustryDGX agent

OpenAI is rolling out the latest version of its AI-powered image generator with new 'thinking capabilities,' allowing it to search the web to help it create multiple images from a single prompt. On Tu

OPeRA: A Dataset of Observation, Persona, Rationale, and Action for Evaluating LLMs on Human Online Shopping Behavior Simulation

Model ReleasesDGX agent

arXiv:2506.05606v5 Announce Type: replace Abstract: Can large language models (LLMs) accurately simulate the next web action of a specific user? While LLMs have shown promising capabilities in generat

Operationalizing Fairness in Text-to-Image Models: A Survey of Bias, Fairness Audits and Mitigation Strategies

SafetyDGX agent

arXiv:2604.16516v1 Announce Type: new Abstract: Text-to-Image (T2I) generation models have been widely adopted across various industries, yet are criticized for frequently exhibiting societal stereoty

OPSDL: On-Policy Self-Distillation for Long-Context Language Models

SafetyDGX agent

arXiv:2604.17535v1 Announce Type: new Abstract: Extending the effective context length of large language models (LLMs) remains a central challenge for real-world applications. While recent post-traini

Optimal control of differentially flat underactuated planar robots in the perspective of oscillation mitigation

ResearchDGX agent

arXiv:2603.15528v2 Announce Type: replace Abstract: Underactuated robots are characterized by a larger number of degrees of freedom than actuators and if they are designed with a specific mass distrib

Optimally Bridging Semantics and Data: Generative Semantic Communication via Schrodinger Bridge

TutorialsDGX agent

arXiv:2604.17802v1 Announce Type: cross Abstract: Generative Semantic Communication (GSC) is a promising solution for image transmission over narrow-band and high-noise channels. However, existing GSC

OptiMVMap: Offline Vectorized Map Construction via Optimal Multi-vehicle Perspectives

Model ReleasesDGX agent

arXiv:2604.17135v1 Announce Type: new Abstract: Offline vectorized maps constitute critical infrastructure for high-precision autonomous driving and mapping services. Existing approaches rely predomin

OptunaHub: A Platform for Black-Box Optimization

Model ReleasesDGX agent

arXiv:2510.02798v2 Announce Type: replace Abstract: Black-box optimization (BBO) underpins advances in domains such as AutoML and Materials Informatics, yet implementations of algorithms and benchmark

Ordering with the Starbucks ChatGPT app was a true coffee nightmare

IndustryDGX agent

Venti iced coffee, light skim milk. That's what I get at Starbucks. It is what I have gotten at Starbucks every time I've been to Starbucks for as long as I can remember, other than a brief love affai

ORSIFlow: Saliency-Guided Rectified Flow for Optical Remote Sensing Salient Object Detection

ResearchDGX agent

arXiv:2603.28584v2 Announce Type: replace Abstract: Optical Remote Sensing Image Salient Object Detection (ORSI-SOD) remains challenging due to complex backgrounds, low contrast, irregular object shap

Our core mission today is using AI to solve document OCR. All of our product offerings, from commercial (LlamaParse) to open-source (LitePar…

AgentsDGX agent

Our core mission today is using AI to solve document OCR. All of our product offerings, from commercial (LlamaParse) to open-source (LiteParse, ParseBench), are fully aligned towards solving this prob

Ouroboros: Single-step Diffusion Models for Cycle-consistent Forward and Inverse Rendering

ResearchDGX agent

arXiv:2508.14461v3 Announce Type: replace Abstract: While multi-step diffusion models have advanced both forward and inverse rendering, existing approaches often treat these problems independently, le

Overcoming Selection Bias in Statistical Studies With Amortized Bayesian Inference

Model ReleasesDGX agent

arXiv:2604.18319v1 Announce Type: cross Abstract: Selection bias arises when the probability that an observation enters a dataset depends on variables related to the quantities of interest, leading to

OVOD-Agent: A Markov-Bandit Framework for Proactive Visual Reasoning and Self-Evolving Detection

SafetyDGX agent

arXiv:2511.21064v2 Announce Type: replace-cross Abstract: Open-Vocabulary Object Detection (OVOD) aims to enable detectors to generalize across categories by leveraging semantic information. Although

P-Check: Advancing Personalized Reward Model via Learning to Generate Dynamic Checklist

ResearchDGX agent

arXiv:2601.02986v2 Announce Type: replace Abstract: Recent approaches in personalized reward modeling have primarily focused on leveraging user interaction history to align model judgments with indivi

PA-TCNet: Pathology-Aware Temporal Calibration with Physiology-Guided Target Refinement for Cross-Subject Motor Imagery EEG Decoding in Stroke Patients

ResearchDGX agent

arXiv:2604.16554v1 Announce Type: new Abstract: Stroke patient cross-subject electroencephalography (EEG) decoding of motor imagery (MI) brain-computer interface (BCI) is essential for motor rehabilit

PAC-Bayes Bounds for Gibbs Posteriors via Singular Learning Theory

Model ReleasesDGX agent

arXiv:2604.17219v1 Announce Type: cross Abstract: We derive explicit non-asymptotic PAC-Bayes generalization bounds for Gibbs posteriors, that is, data-dependent distributions over model parameters ob

Parallel Test-Time Scaling for Latent Reasoning Models

Model ReleasesDGX agent

arXiv:2510.07745v4 Announce Type: replace Abstract: Parallel test-time scaling (TTS) is a pivotal approach for enhancing large language models (LLMs), typically by sampling multiple token-based chains

Parkinson's Disease Detection via Self-Supervised Dual-Channel Cross-Attention on Bilateral Wrist-Worn IMU Signals

ResearchDGX agent

arXiv:2604.18372v1 Announce Type: new Abstract: Parkinson's disease (PD) is a chronic neurodegenerative disease. It shows multiple motor symptoms such as tremor, bradykinesia, postural instability, fr

PARM: Pipeline-Adapted Reward Model

ApplicationsDGX agent

arXiv:2604.18327v1 Announce Type: cross Abstract: Reward models (RMs) are central to aligning large language models (LLMs) with human preferences, powering RLHF and advanced decoding strategies. While

ParseBench is the first benchmark to include VLM chart understanding 📊📈📉 over enterprise documents. 🟠 Existing benchmarks (ChartQA, Char…

Model ReleasesDGX agent

ParseBench is the first benchmark to include VLM chart understanding 📊📈📉 over enterprise documents. 🟠 Existing benchmarks (ChartQA, ChartXiv) test over charts specifically and not the chart's inclusio

PaTaRM: Bridging Pairwise and Pointwise Signals via Preference-Aware Task-Adaptive Reward Modeling

SafetyDGX agent

arXiv:2510.24235v3 Announce Type: replace Abstract: Reward models (RMs) are central to reinforcement learning from human feedback (RLHF), providing the critical supervision signals that align large la

Path-Based Quantum Meta-Learning for Adaptive Optimization of Reconfigurable Intelligent Surfaces

TutorialsDGX agent

arXiv:2604.17690v1 Announce Type: cross Abstract: Reconfigurable intelligent surfaces (RISs) modify signal reflections to enhance wireless communication capabilities. Classical RIS phase optimization

PBSBench: A Multi-Level Vision-Language Framework and Benchmark for Hematopathology Whole Slide Image Interpretation

Model ReleasesDGX agent

arXiv:2604.17570v1 Announce Type: new Abstract: Peripheral Blood Smear (PBS) is a critical microscopic examination in hematopathology that yields whole-slide imaging (WSI). Unlike solid tissue patholo

PCM-NeRF: Probabilistic Camera Modeling for Neural Radiance Fields under Pose Uncertainty

ResearchDGX agent

arXiv:2604.17831v1 Announce Type: new Abstract: Neural surface reconstruction methods typically treat camera poses as fixed values, assuming perfect accuracy from Structure-from-Motion (SfM) systems.

← Previous
1…12381239124012411242…1411
Next →