AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
Human
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,295 results
7 Aug 2026

PoolBench: A Benchmark for Pooling Strategies in Concept Representation Evaluation for Decoder-Only LLMs

Model ReleasesDGX agent

arXiv:2608.05162v1 Announce Type: new Abstract: Pooling is a consequential but under-examined design choice in decoder-only concept representation work: practitioners must collapse token-level hidden

Position: It's Time to Optimize LLMs for Self-Consistency

Model ReleasesDGX agent

arXiv:2608.05188v1 Announce Type: cross Abstract: Despite ever-increasing sophistication in language model (LM) pre- and post-training pipelines, many important failures persist: models overcondition

Positive-Unlabeled Preference Optimization For Chest X-ray Report Generation

ApplicationsDGX agent

arXiv:2608.05341v1 Announce Type: new Abstract: Vision-Language Models (VLMs) for radiology report generation are typically trained on retrospective clinical reports, which suffer from omission noise:

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Post-Hoc Trajectory-Risk Certification for Modular LLM-Based Security Agents

AgentsDGX agent

arXiv:2608.05199v1 Announce Type: cross Abstract: Autonomous security agents operate as staged pipelines, such as classifying network traffic and then attributing attacks to a specific technique. Spli

Posture and Sustainment Optimization Under Adversarial Uncertainty

ResearchDGX agent

arXiv:2608.05256v1 Announce Type: new Abstract: Pre-commitment posture, the assignment of military assets to theater locations before conflict scenarios resolve, is a critical and formally unsolved pr

Potential Matching Optimal Transport: Continuous Normalizing Flows for Exact p-Wasserstein Dynamics

ResearchDGX agent

arXiv:2608.05666v1 Announce Type: new Abstract: We introduce Potential Matching Optimal Transport (PMOT), a potential-flow framework for general p-cost optimal transport with c_p(x,y)=|x-y|^p. PMOT pa

PPDL: LLM-Based Flows as Probabilistic Programs

AgentsDGX agent

arXiv:2608.05234v1 Announce Type: new Abstract: Building reliable applications that leverage large language models (LLMs) remains a significant challenge. While LLMs offer impressive capabilities acro

Predicting Social Media User Actions: A Hybrid Approach for Common and Rare Behavior Prediction on Bluesky

ResearchDGX agent

arXiv:2511.17241v2 Announce Type: replace Abstract: Understanding and predicting user behavior on social media platforms is crucial for content recommendation and platform design. While existing appro

Predicting Task Difficulty Without Rollouts

AgentsDGX agent

arXiv:2608.05797v1 Announce Type: cross Abstract: Task difficulty dictates an agent's likelihood of success, and estimating it without rollouts means forecasting this directly from a task description

Prior-SG: Task and Prior Driven Region Segmentation for Scene Graphs in Arbitrarily-Structured Environments

Local AiDGX agent

arXiv:2608.06170v1 Announce Type: cross Abstract: Hierarchical 3D scene graphs are a promising representation for high-level spatial reasoning in autonomous mobile platforms. However, existing extract

PRISM: Distribution-Gated Flow Matching for Controllable Unpaired Image Translation

Local AiDGX agent

arXiv:2608.06240v1 Announce Type: cross Abstract: Unpaired image-to-image translation must decide, per image, what to change and what to preserve without paired supervision. Many diffusion-based unpai

PRISM: Priority-aware Rubric Internalization via Structured Multimodal Data Synthesis

ApplicationsDGX agent

arXiv:2608.05249v1 Announce Type: cross Abstract: Real-world multimodal instructions often bundle multiple requirements with unequal importance, yet most multimodal training data still reduce instruct

ProDVI: Programmatic Dynamics Priors for Value Network Initialization

ResearchDGX agent

arXiv:2608.06015v1 Announce Type: cross Abstract: Deep Reinforcement Learning (RL) is notoriously sample inefficient. One contributing factor is that RL agents are typically initialized from scratch,

Project2Task: Graph-Guided Project-Level Planning for Autonomous Research

Model ReleasesDGX agent

arXiv:2608.05225v1 Announce Type: new Abstract: Research agents can increasingly search literature, propose hypotheses, generate code, run experiments, and draft manuscripts from a single topic. Howev

PromptForSegCXR: Prompt-Driven Multi-Organ and Multi-Disease Segmentation in Chest X-rays using a Multi-stage Fusion Mechanism

ResearchDGX agent

arXiv:2507.00673v2 Announce Type: replace-cross Abstract: Image segmentation is central to automated medical image analysis, enabling precise identification of anatomical structures and pathological r

Provably Efficient Self-Calibrating Quantum Fault Tolerance

ResearchDGX agent

arXiv:2608.05686v1 Announce Type: cross Abstract: Quantum error correction protects logical information only when every physical operation remains below the fault-tolerance threshold, a condition that

QEvict: Recoverable Quantized KV Eviction for Attention-Drift-Robust Long-Context Decoding

ResearchDGX agent

arXiv:2608.05326v1 Announce Type: cross Abstract: Autoregressive large language model inference is increasingly constrained by the memory footprint of the Key-Value (KV) cache. A dominant line of work

Quality Diversity for Reliable Data Driven Time-Use Optimization

ResearchDGX agent

arXiv:2608.05230v1 Announce Type: cross Abstract: The daily allocation of the finite 24-hour time budget is strongly associated with physical, mental, and cognitive health. While predictive models can

QuanTiMedAI: Quantum-Enhanced Time-Series Model guided by Agentic AI for Cardiac Arrest Mortality Prediction

AgentsDGX agent

arXiv:2608.06294v1 Announce Type: new Abstract: Cardiac arrest remains one of the most lethal conditions encountered in intensive care units. Despite the growing availability of electronic health reco

Quantum-Structured World Models (QSWMs) for Predictive Latent Dynamics

TutorialsDGX agent

arXiv:2608.05371v1 Announce Type: new Abstract: World models learn latent states that summarize interaction histories, evolve over time, and support prediction, simulation, or planning. Most existing

RA-CAD: Learning Post-Execution Critique for State-Aware Text-to-CAD Generation

SafetyDGX agent

arXiv:2608.05714v1 Announce Type: new Abstract: Text-to-CAD generation translates natural-language design intent into editable and executable parametric computer-aided design (CAD) codes, reducing the

RASP-QAOA: Resource-Aware Per-Instance Selection for Exact QAOA Simulation

SafetyDGX agent

arXiv:2608.05646v1 Announce Type: cross Abstract: Exact QAOA simulation spans several computational representations whose useful regions differ sharply across graph structure, circuit depth, precision

RealityBridge: Bridging Editable 3D Gaussian Splatting Driving Simulations and Real-World Videos

Local AiDGX agent

arXiv:2606.16278v3 Announce Type: replace-cross Abstract: Long-tail hazardous scenarios are essential for safety-oriented autonomous driving, yet they are difficult to collect at scale. Editable 3D Ga

Reasoning Errors Have a Region and a Direction in the Residual-Stream Trajectory of LLMs

ResearchDGX agent

arXiv:2608.05660v1 Announce Type: cross Abstract: As language models are increasingly used for tasks that require verifiable reasoning, reliably distinguishing sound reasoning from flawed reasoning ha

Rectifying Geometric Misalignment: Online Source-Free Adaptation for Class-Imbalanced EEG

Model ReleasesDGX agent

arXiv:2608.05315v1 Announce Type: new Abstract: Electroencephalography (EEG) based Brain-Computer Interfaces (BCIs) often require unsupervised domain adaptation (UDA) to generalize across subjects and

Recursive Synthesis for Long-Horizon Terminal Tasks

Model ReleasesDGX agent

arXiv:2608.05466v1 Announce Type: new Abstract: High-quality long-horizon training data for terminal agents is expensive to produce, often costing hundreds to thousands of dollars per task, because ea

Reducing belief in conspiracy theories as they unfold using large language models

ResearchDGX agent

arXiv:2608.06151v1 Announce Type: cross Abstract: The emergence of conspiracy theories in the wake of major events is a significant societal challenge. Here we test whether conversational dialogues wi

Refining Over Resampling: Test-Time Self-Correction for LLM Reasoning

ResearchDGX agent

arXiv:2608.05643v1 Announce Type: new Abstract: Test-time scaling improves LLM reasoning by using additional inference compute, but wider sampling alone can suffer from diminishing returns: new rollou

Reinforcing Action Policies by Prophesying

TutorialsDGX agent

arXiv:2511.20633v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) policies excel in aligning language, perception, and robot control. However, most VLAs are trained purely by imitation,

Relay, Don't Route: Adaptive Population Handoff for Cost-Efficient LLM-Driven Evolution

ResearchDGX agent

arXiv:2608.05651v1 Announce Type: cross Abstract: Large language model (LLM)-driven evolution has shown promise for program search and algorithm discovery, but relying on strong models throughout long

Resourced Authority A Mechanism-Design Model for Participatory Governance of Deployed AI Agents

SafetyDGX agent

arXiv:2608.06353v1 Announce Type: cross Abstract: We give a formal mechanism design model for the continuous participatory governance of a deployed AI agent. The mechanism is built on the principle th

Respect Your Zero-Shot Uncertainty: Conservative Calibration for Test-Time-Adapted Vision-Language Models

ResearchDGX agent

arXiv:2608.05945v1 Announce Type: new Abstract: Test-time adaptation (TTA) can improve the recognition accuracy of vision-language models under distribution shift, but often degrades calibration, maki

Reversible Unlearnable Examples: Towards the Copyright Protection in Deep Learning Era

TutorialsDGX agent

arXiv:2608.06211v1 Announce Type: cross Abstract: Significant advancements in deep learning have been made possible by the utilization of large datasets, underscoring the critical importance of copyri

Revisiting Black-Box Model Ownership Verification through Information Theory

ApplicationsDGX agent

arXiv:2409.06130v2 Announce Type: replace-cross Abstract: Modern machine learning models require substantial computational resources and data to train, making them valuable intellectual property. Mode

RIG-RoPE: Relation- and Instance-Gated Rotary Positional Encoding with Duration-Aware Temporal Coordinates

ResearchDGX agent

arXiv:2608.05154v1 Announce Type: new Abstract: Rotary positional encoding (RoPE) is a core component of modern language models and has been extended to multimodal LLMs through multidimensional varian

Robot Learning from Human Demonstrations: Handwritten Alphabet Trajectories and Human-Likeness Evaluation

Model ReleasesDGX agent

arXiv:2608.06221v1 Announce Type: cross Abstract: Learning from demonstration (LfD) provides a developmental framework through which robots can develop motor skills by observing and imitating human dy

Robust Context-Aware Detection of Malicious Instructions in Text

AgentsDGX agent

arXiv:2608.05430v1 Announce Type: cross Abstract: The remarkable instruction-following ability of modern LLMs has enabled their practical use as the minds of agents that can autonomously complete incr

Robust Native Language Identification through Agentic Decomposition

Model ReleasesDGX agent

arXiv:2509.16666v2 Announce Type: replace Abstract: Large language models (LLMs) often achieve high performance in native language identification (NLI) benchmarks by leveraging superficial contextual

Robust-WAM: Bridging Generative Pretraining and Semantic Foresight in World-Action Models

SafetyDGX agent

arXiv:2608.05903v1 Announce Type: new Abstract: Mainstream World-Action Models (WAMs) adapt pretrained video generation models (VGMs) for robot control, transferring their learned dynamics prior for a

Routing Is Least Learnable Where It Is Most Valuable: Bounds on Representation Routing for Web Agents

AgentsDGX agent

arXiv:2608.06171v1 Announce Type: new Abstract: Web agents observe a browser through text, pixels, or both, and the choice is usually fixed once for all tasks. We measure six observation modes across

RP-OPSD: Reasoning-Pivot-Guided On-Policy Self-Distillation for Multilingual Reasoning Transfer

SafetyDGX agent

arXiv:2608.06347v1 Announce Type: new Abstract: Multilingual reasoning transfer is crucial for extending reasoning capabilities of large language models (LLMs) beyond high-resource languages. On-polic

RRC: Unlocking Generative Reward Models in LLM Reinforcement Learning via Ranking-Based Reward Construction

ResearchDGX agent

arXiv:2608.06310v1 Announce Type: cross Abstract: Recent advances in reward modeling show a paradigm shift from discriminative reward models to generative reward models. However, despite their strong

Runtime Observability for Heterogeneous Attention Memory

Model ReleasesDGX agent

arXiv:2608.05863v1 Announce Type: new Abstract: Modern models no longer keep a plain KV cache: latent caches, learned sparse selectors and recurrent states each carry the model's memory in a different

RxnCLF: Contrastive Transformation-Aware Reaction Foundation Model for Improved Reactivity Prediction

TutorialsDGX agent

arXiv:2608.06259v1 Announce Type: new Abstract: Reaction yield prediction remains challenging because labeled data are scarce and reaction space is both combinatorially large and sparsely populated, l

Safe Evolution with Circuit Anchors

SafetyDGX agent

arXiv:2608.05158v1 Announce Type: new Abstract: In biological evolution, unconstrained mutation can lead to catastrophic outcomes: organisms may evolve enhanced capabilities while losing essential fun

SafeDivertor: Faithful Divertor Heat Flux Reconstruction from Macroscopic Plasma State Signals via Time-Frequency Prior Exploitation

Model ReleasesDGX agent

arXiv:2608.05669v1 Announce Type: cross Abstract: Divertor heat-flux analysis is essential for understanding plasma-wall interactions and protecting plasma-facing components in magnetic-confinement fu

SAGA: Score-Weighted Adaptive Generation Alignment for Low-Resource Nordic Language Models

SafetyDGX agent

arXiv:2608.06179v1 Announce Type: new Abstract: Preference optimisation has proven effective for improving large language models but typically relies on costly human preference annotations. Extending

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training

SafetyDGX agent

arXiv:2608.06125v1 Announce Type: new Abstract: Latent reward models can supervise visual diffusion models without decoding intermediate states into pixel space. This makes alignment with human prefer

Scaffold-Mediated Post-Training: Co-Evolving Model Parameters and Procedural Scaffold Graphs

Model ReleasesDGX agent

arXiv:2608.05156v1 Announce Type: new Abstract: Post-training of large language models optimizes only parameters, while inference-time procedural scaffolds are typically designed independently of para

Scalable estimation of VARMA models

SafetyDGX agent

arXiv:2608.06340v1 Announce Type: cross Abstract: Vector autoregressive moving-average (VARMA) models have long been considered impractical beyond moderate dimensions: the likelihood is non-convex, th

Schema-Guided Hierarchical Information Extraction and Semantic Evaluation Using Generative AI

Model ReleasesDGX agent

arXiv:2608.06167v1 Announce Type: new Abstract: We present a schema-based framework for extracting complex, structured information from unstructured text documents using generative AI, followed by aut

SCI-CLIP: Segment-Centric Inference with Reference Memory for Training-Free Open-Vocabulary Segmentation

SafetyDGX agent

arXiv:2608.05627v1 Announce Type: new Abstract: Training-free open-vocabulary segmentation remains limited by a missing inference abstraction. Frozen vision-language features are produced at patch lev

Scientific Machine Learning of Chaotic Systems Learns Reduced-Order Equations for Neural Populations

Model ReleasesDGX agent

arXiv:2507.03631v4 Announce Type: replace Abstract: Extracting interpretable mathematical models from complex dynamical systems is difficult, especially for chaotic dynamics observed with noisy experi

SciQNet: Two-Stage Multimodal Adaptation for Scientific Image Quality Assessment

ResearchDGX agent

arXiv:2608.05691v1 Announce Type: new Abstract: Scientific images are essential for communicating experimental observations, quantitative evidence and conceptual knowledge. Unlike natural images, thei

SCP-NL2TL: Selective Conformal Prediction with Semantic Verification for Natural Language to Temporal Logic Specifications

SafetyDGX agent

arXiv:2608.05439v1 Announce Type: new Abstract: Translating natural language instructions into machine-interpretable formal specifications enables robots and autonomous systems to plan, reason, and fo

SEAM: Global consistency beyond local accuracy in scientific machine learning

Model ReleasesDGX agent

arXiv:2608.05702v1 Announce Type: new Abstract: Scientific machine learning commonly validates models at the level of a subdomain, a benchmark split, or an explanation for one prediction. Yet such loc

Search-Aided Joint Agent-Environment Reinforcement Learning for Robust Lifelong Multi-Agent Path Finding with Rotations

SafetyDGX agent

arXiv:2608.05588v1 Announce Type: cross Abstract: Lifelong Multi-Agent Path Finding (LMAPF) requires repeatedly planning collision-free paths for agents that continuously receive new goals upon reachi

Search2Skill: Skill Distillation Beyond Knowledge Boundaries Via Rubric-Based Reinforcement Learning

AgentsDGX agent

arXiv:2608.05245v1 Announce Type: new Abstract: Reusable skills, which encapsulate the procedural knowledge required to solve real-world professional tasks, offer LLM-based agents a path toward self-e

SearchAuditor: Auditing and Attributing Failures in Long-Horizon Search Agents

Model ReleasesDGX agent

arXiv:2608.05212v1 Announce Type: new Abstract: Deep search agents tackle challenging questions through long-horizon web interactions, a process that is both complex and fragile: small reasoning error

Seeing Is Not Deciding: Can Multimodal LLMs Act as Effective CEOs?

Model ReleasesDGX agent

arXiv:2608.05864v1 Announce Type: new Abstract: Large language models are increasingly applied as autonomous decision-making agents. However, in executive business decisions, existing benchmarks are l

← Previous
1…5455565758…989
Next →