AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,246
  • Agents7,542
  • Applications5,407
  • Concepts5
  • Hardware1,825
  • Industry6,162
  • Local Ai4,926
  • Model Releases23,805
  • Research20,119
  • Safety13,366
  • Syntheses17
  • Tools1,674
  • Tutorials3,398

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,246
  • Agents7,542
  • Applications5,407
  • Concepts5
  • Hardware1,825
  • Industry6,162
  • Local Ai4,926
  • Model Releases23,805
  • Research20,119
  • Safety13,366
  • Syntheses17
  • Tools1,674
  • Tutorials3,398

Source
Human
88,246Total entries
1Added by human
88,245Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,897 results
29 May 2026

On Language Generation in the Limit with Bounded Memory

ResearchDGX agent

arXiv:2605.30324v1 Announce Type: cross Abstract: We study language generation in the limit under bounded memory. In this task, a learner observes examples from an unknown target language one at a tim

On-Policy Replay for Continual Supervised Fine-Tuning

Model ReleasesDGX agent

arXiv:2605.29495v1 Announce Type: new Abstract: Continual supervised fine-tuning (SFT) is the de facto recipe for adapting large language models (LLMs) to a stream of downstream tasks, but it suffers

On the Construction and Implications of Low-Loss Valleys in LoRA-based Bayesian Inference

Model ReleasesDGX agent

arXiv:2605.29580v1 Announce Type: new Abstract: While parameter-efficient fine-tuning methods like low-rank adaptation (LoRA) are standard for large language models, principled estimation of epistemic

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

On the Geometry of Games and their Solvers

SafetyDGX agent

arXiv:2605.29919v1 Announce Type: new Abstract: A central challenge in game theory and learning systems such as GANs is understanding which algorithms can efficiently compute equilibria across the het

On the Optimizer Dependence of Neural Scaling Laws

ResearchDGX agent

arXiv:2605.29387v1 Announce Type: cross Abstract: The scaling exponent alpha in neural scaling laws L(N) propto N^{-alpha} is commonly treated as a fixed constant set by architecture and data. We pres

One Click per Cell Type Suffices: Training-free Group Interaction for Cell Instance Segmentation

ResearchDGX agent

arXiv:2605.29429v1 Announce Type: new Abstract: Cell instance segmentation models trained on cell-specific datasets suffer severe performance drops on out-of-distribution cell types, while interactive

One Mask to Rule Them All: On Hidden Facts after Editing and How to Find Them

TutorialsDGX agent

arXiv:2605.28839v1 Announce Type: new Abstract: Knowledge editing methods such as ROME and MEMIT update factual associations in transformer models by modifying MLP weights. While evaluated mainly by o

Online Fair Division with Additional Information

SafetyDGX agent

arXiv:2505.24503v3 Announce Type: replace-cross Abstract: We study the problem of fairly allocating indivisible goods to agents in an online setting, where goods arrive sequentially and must be alloca

OOD-GraphLLM: Graph Large Language Model for Out-of-Distribution Generalized Drug Synergy Prediction

Model ReleasesDGX agent

arXiv:2605.30247v1 Announce Type: new Abstract: Drug synergy prediction (DSP) aims to identify efficacious drug combinations under various cellular contexts with different targets. However, the contin

Open Problem: Separating Geometric and Algorithmic Compression via Cayley-Table Completion

SafetyDGX agent

arXiv:2605.29885v1 Announce Type: new Abstract: Modern statistical learning theory and deep learning characterize generalization primarily in terms of continuous capacity control (e.g., norm-based reg

Open World Autoencoding Drift Detection with Novel Class Recognition in Tabular Non-stationary Data Streams

ResearchDGX agent

arXiv:2605.29834v1 Announce Type: new Abstract: Data stream processing has become a landmark in modern machine learning applications, with concept drifts and novel class appearances posing the primary

OpenClawBench: Benchmarking Process-side Anomalies in Real-world Agent Execution Trajectories

Model ReleasesDGX agent

arXiv:2605.29253v1 Announce Type: new Abstract: Task success can hide process anomalies in real-world agent executions. An agent may pass the final task oracle while still accumulating unresolved ambi

Opir: Efficient Multi-Task Safety Classification for Toxicity, Jailbreaks, Hate Speech, and Harmful Content

Model ReleasesDGX agent

arXiv:2605.29659v1 Announce Type: cross Abstract: Real-time safety filtering for large language model (LLM) applications requires classifiers that can detect unsafe prompts, toxic language, jailbreak

Opt-Verifier: Unleashing the Power of LLMs for Optimization Modeling via Dual-Side Verification

ResearchDGX agent

arXiv:2605.29556v1 Announce Type: new Abstract: Building mathematical optimization models is critical in operations research (OR), while it requires substantial human expertise. Recent advancements ha

Optimal Gap-Dependent Regret for Private Stochastic Decision-Theoretic Online Learning

ResearchDGX agent

arXiv:2605.29148v1 Announce Type: new Abstract: We study stochastic decision-theoretic online learning with full information and event-level pure differential privacy. A COLT open problem of Hu and Me

Optimal Rates for Differentially Private Hypothesis Testing with E-values

ResearchDGX agent

arXiv:2605.28952v1 Announce Type: cross Abstract: E-values have attracted considerable interest in recent years as flexible tools for enabling anytime-valid and adaptive data analysis. Hypothesis test

Optimization and Generation in Aerodynamics Inverse Design

ResearchDGX agent

arXiv:2602.03582v3 Announce Type: replace Abstract: Aerodynamic inverse design can improve vehicle and aircraft efficiency, but practical design rarely seeks performance alone: vehicle refinement must

Optimizing Latent Representations for Robust Building Damage Assessment Onboard Earth Observation Satellites

Model ReleasesDGX agent

arXiv:2605.29575v1 Announce Type: new Abstract: Rapid identification of damaged buildings after natural disasters or on war areas is crucial to support emergency response and prioritize interventions.

OptSkills: Learning Generalizable Optimization Skills from Problem Archetypes via Cluster-Based Distillation

Model ReleasesDGX agent

arXiv:2605.29829v1 Announce Type: new Abstract: Leveraging Large Language Models (LLMs) to automatically formulate and solve optimization problems from natural language has emerged as an efficient par

Order-Agnostic Autoregressive Modelling with Missing Data

ApplicationsDGX agent

arXiv:2605.06355v2 Announce Type: replace Abstract: Order-Agnostic autoregressive models have demonstrated strong performance in deep generative modeling, yet their use in settings with incomplete dat

Orthogonal Concept Erasure for Diffusion Models

Model ReleasesDGX agent

arXiv:2605.28902v1 Announce Type: new Abstract: Concept erasure has emerged as a promising approach to mitigate undesired or unsafe content in diffusion models, yet existing methods still face signifi

Orthogonal Negative Guidance in Attention Feature Space for Text-to-Image Generation

SafetyDGX agent

arXiv:2605.29390v1 Announce Type: new Abstract: Text-to-image (T2I) models have become increasingly capable of generating high-quality images. Yet, enforcing the explicit absence of a specified object

OVA-IB: One vs All Information Bottleneck for Multi-Modal Alignment

Model ReleasesDGX agent

arXiv:2605.29900v1 Announce Type: new Abstract: Contrastive learning is effective for aligning paired views or modalities, but alignment beyond two modalities remains non-trivial and comparatively und

Overcoming Forgetting in LLM Fine-Tuning with Evolution Strategies

Model ReleasesDGX agent

arXiv:2605.30148v1 Announce Type: cross Abstract: Evolution Strategies (ES) has recently emerged as a competitive alternative to reinforcement learning (RL) for large language model (LLM) fine-tuning,

P^2RAG: Efficient Privacy-Preserving RAG Service Supporting Arbitrary Top-k Retrieval

ApplicationsDGX agent

arXiv:2603.14778v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) enables large language models to use external knowledge, but outsourcing the RAG service raises privacy c

Paper Agents, Paper Gains: An Empirical Analysis of DeFi Investment Agents

SafetyDGX agent

arXiv:2605.29174v1 Announce Type: new Abstract: DeFi investment agents, systems that use AI for autonomous on-chain trading, have attained over USD 3 billion in combined token valuations since late 20

Parallax: Parameterized Local Linear Attention for Language Modeling

Model ReleasesDGX agent

arXiv:2605.29157v1 Announce Type: cross Abstract: Large Language Models (LLMs) have become the central paradigm in artificial intelligence, yet the core computational primitive of attention has remain

Parallel Adaptive Multi-Objective Evolutionary Learning of Discretized Bayesian Network Classifiers for Clinical Data

ApplicationsDGX agent

arXiv:2605.29058v1 Announce Type: new Abstract: Bayesian Networks (BNs) are of interest from an explainable AI viewpoint, offering transparent probabilistic models for decision support. Baymex is a re

Parameter-Efficient Subspace Decoupling ViT for Mitigating Multi-Task Negative Transfer in Histological Scoring

Model ReleasesDGX agent

arXiv:2605.29852v1 Announce Type: new Abstract: Histological scoring is essential for diagnosing Non-Alcoholic Fatty Liver Disease (NAFLD), yet its automation remains challenging due to the high annot

ParaTool: Shifting Tool Representations from Context to Parameters

Model ReleasesDGX agent

arXiv:2605.29561v1 Announce Type: new Abstract: Tool calling extends large language models (LLMs) by enabling grounded interaction with external executable interfaces, thereby supporting environment-c

PARCEL: Pool-Anchored Resampling with Conditioned Elastic Queries for Efficient Vision-Language Understanding

ApplicationsDGX agent

arXiv:2605.30126v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) map visual inputs into dense token sequences, imposing a quadratic computational bottleneck for inference. Elasti

ParCo-SDF: Learning Prior-Free Partial-to-Complete Signed Distance Fields of Deformable Objects

ResearchDGX agent

arXiv:2605.29417v1 Announce Type: new Abstract: This study addresses the partial-to-complete geometry reconstruction of deformable objects (DOs) from point-cloud observations toward precise DO manipul

PassNet: Scaling Large Language Models for Graph Compiler Pass Generation

ApplicationsDGX agent

arXiv:2605.29357v1 Announce Type: new Abstract: Modern tensor compilers such as TorchInductor deliver substantial speedups on mainstream models, yet face a systematic performance ceiling on long-tail

PatchBoard: Schema-Grounded State Mutation for Reliable and Auditable LLM Multi-Agent Collaboration

AgentsDGX agent

arXiv:2605.29313v1 Announce Type: new Abstract: LLM multi-agent systems often coordinate through natural-language dialogue or loosely structured shared memory, making intermediate state difficult to v

Path-Space Mirror Descent for On-Policy Reinforcement Learning under the Generalized Schrodinger Bridge

SafetyDGX agent

arXiv:2603.21621v2 Announce Type: replace Abstract: Classical on-policy algorithms such as PPO and mirror descent policy optimization provide stable proximal policy updates through tractable action li

PEARL: Training Socratic Tutors with Pedagogically Aligned Reinforcement Learning

SafetyDGX agent

arXiv:2605.29582v1 Announce Type: cross Abstract: Large Language Models (LLMs) have shown promise as educational tutors, yet effective tutoring requires more than solving problems: it must provide pro

Permutation-Invariant Spectral Learning via Dyson Diffusion

SafetyDGX agent

arXiv:2510.08535v2 Announce Type: replace-cross Abstract: Diffusion models are central to generative modeling and have been adapted to graphs by diffusing adjacency matrix representations. The challen

Persona Conditioning of Brand Recommendations in Retrieval-Augmented Commercial Chat: A Prominence-Stratified Cross-Provider Audit

ApplicationsDGX agent

arXiv:2605.30207v1 Announce Type: new Abstract: The same prompt -- 'best CRM software' -- reaches AI assistants from buyers in widely different contexts: a solo founder, an enterprise VP, a UK SMB own

PersonaAgent: Bridging Memory and Action for Personalized LLM Agents

SafetyDGX agent

arXiv:2506.06254v2 Announce Type: replace Abstract: Large Language Model (LLM) empowered agents have recently emerged as advanced paradigms that exhibit impressive capabilities in a wide range of doma

Personalized Turn-Level User Conversation Satisfaction Benchmark

Model ReleasesDGX agent

arXiv:2605.29711v1 Announce Type: cross Abstract: User satisfaction with AI assistants is highly personalized: the same response may satisfy one user but disappoint another depending on what each user

PhAIL: A Real-Robot VLA Benchmark and Distributional Methodology

Model ReleasesDGX agent

arXiv:2605.29710v1 Announce Type: new Abstract: Real-world evaluation of vision-language-action (VLA) policies still rests on binary success rate at a fixed timeout with N le 25 rollouts per condition

Phantom: Training Robots Without Robots Using Only Human Videos

ResearchDGX agent

arXiv:2503.00779v2 Announce Type: replace Abstract: Training general-purpose robots requires learning from large and diverse data sources. Current approaches rely heavily on teleoperated demonstration

Phase-Conditioned Imitation Learning with Autonomous Failure Recovery for Robust Deformable Object Manipulation

SafetyDGX agent

arXiv:2605.29407v1 Announce Type: new Abstract: This paper presents a phase-conditioned, force-aware framework for robust deformable object manipulation. Standard imitation learning policies such as A

PhoneWorld: Scaling Phone-Use Agent Environments

Model ReleasesDGX agent

arXiv:2605.29486v1 Announce Type: cross Abstract: A central bottleneck for phone-use agents is that controllable, reproducible environments covering real mobile behavior are hard to build at scale. Ex

PhyGenHOI: Physically-Aware 4D Generation of Dynamic Human-Object Interactions

AgentsDGX agent

arXiv:2605.30268v1 Announce Type: cross Abstract: We address the task of generating physically accurate and visually faithful 4D Human-Object Interaction (HOI). Given a static 3D human and target obje

Physics Is All You Need? A Case Study in Physicist-Supervised AI Development of Scientific Software

Model ReleasesDGX agent

arXiv:2605.30353v1 Announce Type: new Abstract: Are AI agents tools, co-authors, or researchers? We present a quantified case study (N=1): a physicist supervising an AI coding agent (Claude Code, Sonn

Plan, Don't Pose: Long Composite Motion Generation with Text-Aligned BFM

SafetyDGX agent

arXiv:2605.29906v1 Announce Type: new Abstract: Text-to-motion (T2M) generation has broad applications in character animation, virtual avatars, and human-robot interaction. Existing methods typically

Planning with the Views via Scene Self-Exploration

Model ReleasesDGX agent

arXiv:2605.29563v1 Announce Type: new Abstract: Can VLMs predict how each camera move changes the view, and plan many such moves ahead? We call this capability view planning, requiring (1)understandin

Pocket-Dentist: On-Device Dental Image Understanding via Efficient Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2605.29299v1 Announce Type: cross Abstract: Evaluations of dental vision-language models remain fragmented across datasets, task definitions and metrics, and often ignore their computational cos

PokerSkill: LLMs Can Play Expert-Level Poker without Training or Solvers

Model ReleasesDGX agent

arXiv:2605.30094v1 Announce Type: new Abstract: Poker is a landmark challenge for artificial intelligence. The dominant approach relies on equilibrium solvers built on counterfactual regret minimizati

Position: Stop Chasing the C-index when Evaluating Survival Analysis Models

SafetyDGX agent

arXiv:2506.02075v2 Announce Type: replace-cross Abstract: The current state of evaluation in survival analysis is plagued by the persistent use of evaluation metrics in ways that are misaligned with t

Position: Text Embeddings Should Capture Implicit Semantics, Not Just Surface Meaning

ApplicationsDGX agent

arXiv:2506.08354v2 Announce Type: replace-cross Abstract: This position paper argues that text embedding research should move beyond surface meaning and embrace implicit semantics as a central modelin

Practical Insights on Grasp Strategies for Mobile Manipulation in the Wild

TutorialsDGX agent

arXiv:2504.12512v2 Announce Type: replace Abstract: Mobile manipulation robots are continuously advancing, with their grasping capabilities rapidly progressing. However, there are still significant ga

Practitioner Beliefs and Behaviors in AI-Enhanced Education: DOT Framework Survey Evidence

SafetyDGX agent

arXiv:2605.29041v1 Announce Type: new Abstract: This study reports findings from a cross-sectional survey (n = 72) of higher education practitioners examining beliefs, behaviors, and institutional con

PRAIB: Peer Review AI Benchmark of Behaviour of LLM-Assisted Reviewing

Model ReleasesDGX agent

arXiv:2605.29815v1 Announce Type: new Abstract: The growing number of submitted papers has motivated the exploration of Large Language Models (LLMs) as a means to support and augment the peer review p

Pre-Registering the Detectable Effect: A Paired-MDE Budget for 4-bit Quantization Benchmarks, with a Pilot Audit

Model ReleasesDGX agent

arXiv:2605.28873v1 Announce Type: new Abstract: This is a planning-method note with an unpaired pilot audit. We adapt the classical paired-binary sample-size calculation (Miettinen, 1968) to quantizat

Predicting Causal Effects from Natural Language Queries using Structured Representations

Model ReleasesDGX agent

arXiv:2605.29631v1 Announce Type: cross Abstract: Randomized controlled trials are a cornerstone of medicine and the social sciences as they enable reliable estimates of causal effects. However, they

Prediction-Powered Inference Across Many Tasks for AI Evaluation & Social Science Research

ApplicationsDGX agent

arXiv:2605.29249v1 Announce Type: cross Abstract: Many applications require statistically valid inference across many related tasks, while using only a handful of high-quality labels per hypothesis. I

Prescribe-then-Select: Adaptive Policy Selection for Contextual Stochastic Optimization

Model ReleasesDGX agent

arXiv:2509.08194v2 Announce Type: replace Abstract: We address the problem of policy selection in contextual stochastic optimization (CSO), where covariates are available as contextual information and

Prioritize the Process, Not Just the Outcome: Rewarding Latent Thought Trajectories Improves Reasoning in Looped Language Models

Model ReleasesDGX agent

arXiv:2602.10520v3 Announce Type: replace Abstract: Looped Language Models (LoopLMs) perform multi-step latent reasoning prior to token generation and outperform conventional LLMs on reasoning benchma

← Previous
1…560561562563564…1049
Next →