AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
Human
87,678Total entries
1Added by human
87,677Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,383 results
27 May 2026

On the Push-Based Asynchronous Federated Learning: A Bias-Correction Aggregation Approach

SafetyDGX agent

arXiv:2605.26162v1 Announce Type: cross Abstract: Asynchronous decentralized federated learning (ADFL) eliminates central coordination and global synchronization, making it attractive for large-scale

On the Robustness of Machine Unlearning for Vision-Language Models

ResearchDGX agent

arXiv:2605.26992v1 Announce Type: new Abstract: Vision-language models (VLMs) may memorize undesirable information from training data, motivating growing interest in machine unlearning. In this work,

On the Role of Inductive Bias in Time-Series Pretraining: A Case Study in Learning Generalizable Representations for Clinical Time Series

SafetyDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.26194v1 Announce Type: new Abstract: Clinical time-series learning is routinely constrained by small, heterogeneous cohorts and protocol drift, while its downstream use spans both classific

On the Sensitivity of Instruction-tuned LLMs to Harmful Sentences in Long Inputs

Model ReleasesDGX agent

arXiv:2510.05864v2 Announce Type: replace Abstract: Large language models (LLMs) increasingly operate on long inputs, yet their behavior when harmful sentences are sparsely embedded within such inputs

Once-For-All: A Train-Once and Select-Anytime Framework for Multimodal Instruction Tuning

ResearchDGX agent

arXiv:2605.26761v1 Announce Type: new Abstract: Multimodal instruction tuning is the de facto recipe for adapting vision language models (VLMs), yet instruction data are highly redundant, making data

Online Learning on Hidden-Convex Losses via Algorithmic Equivalence: Optimal Regret, Geometric Barrier, and Bandit Feedback

ResearchDGX agent

arXiv:2605.26373v1 Announce Type: new Abstract: We study adversarial online learning with hidden-convex losses, i.e., nonconvex losses that become convex after a nonlinear reparameterization. Ghai, Lu

Open-Weight LLM Fine-Tuning Defenses are Susceptible to Simple Attacks

SafetyDGX agent

arXiv:2605.26526v1 Announce Type: new Abstract: Recent defenses for safeguarding open-weight large language models (LLMs) are intended to prevent adversarial usage. Underlying these defenses is an ass

Optimal Rates for Feasible Payoff Set Estimation in Games

AgentsDGX agent

arXiv:2602.04397v2 Announce Type: replace-cross Abstract: We study a setting in which two players play a (possibly approximate) Nash equilibrium of a bimatrix game, while a learner observes only their

Optimising Factual Consistency in Summarisation via Preference Learning from Multiple Imperfect Metrics

TutorialsDGX agent

arXiv:2605.26840v1 Announce Type: new Abstract: Reinforcement learning with evaluation metrics as rewards is widely used to enhance specific capabilities of language models. However, for tasks such as

ORCA: An End-to-End Interactive Copilot for Optimized Root Cause Analysis

TutorialsDGX agent

arXiv:2605.27022v1 Announce Type: new Abstract: Causal analysis is a crucial task in many domains, including manufacturing, social science, and medicine. However, despite recent progress, the conceptu

ORLoopBench: Solver-in-the-Loop Benchmarks for Self-Correction and Behavioral Rationality in Operations Research

Model ReleasesDGX agent

arXiv:2601.21008v3 Announce Type: replace-cross Abstract: Operations Research practitioners debug infeasible models through an iterative process: inspecting Irreducible Infeasible Subsystems ( IIS), i

OSMa-Bench++: Toward Open-Ended Benchmarking of Semantic Mapping for Manipulation with Prompt-Generated Synthetic Scenes

Model ReleasesDGX agent

arXiv:2605.26831v1 Announce Type: new Abstract: Semantic mapping methods are increasingly used as intermediate scene representations for downstream robotic reasoning and manipulation, yet their evalua

Over-Alignment vs Over-Fitting: The Role of Feature Learning Strength in Generalization

SafetyDGX agent

arXiv:2602.00827v2 Announce Type: replace Abstract: Feature learning strength (FLS), i.e., the inverse of the effective output scaling of a model, plays a critical role in shaping the optimization dyn

Pair-In, Pair-Out: Latent Multi-Token Prediction for Efficient LLMs

SafetyDGX agent

arXiv:2605.27255v1 Announce Type: cross Abstract: Long chain-of-thought reasoning has made autoregressive decoding the dominant inference cost of modern large language models. Existing methods target

PARE: Pruning and Adaptive Routing for Efficient Video Generation

ResearchDGX agent

arXiv:2605.27336v1 Announce Type: new Abstract: Video Diffusion Transformers (DiTs) generate high-quality videos but demand substantial compute due to wide blocks, deep architectures, and iterative sa

Parsimonious Learning-Augmented Online Metric Matching

ResearchDGX agent

arXiv:2605.26886v1 Announce Type: cross Abstract: Learning-augmented algorithms have received significant attention in recent years, particularly in the context of online optimization. Motivated by th

ParsVoice: A Large-Scale Multi-Speaker Persian Speech Corpus for Text-to-Speech Synthesis

ResearchDGX agent

arXiv:2510.10774v3 Announce Type: replace-cross Abstract: Persian remains substantially underrepresented in open speech-text resources, limiting progress in multi-speaker text-to-speech (TTS), speech-

Particle-Lund Multimodality in Jet Taggers

ResearchDGX agent

arXiv:2605.26821v1 Announce Type: cross Abstract: The Lund plane offers a physics-motivated, hierarchical representation of QCD radiation within jets, while transformer-based taggers have reached stat

PashtoTTS-Bench: automated screening for low-resource non-Latin-script text-to-speech

Model ReleasesDGX agent

arXiv:2605.26978v1 Announce Type: new Abstract: Text-to-speech (TTS) evaluation for low-resource non-Latin-script languages can fail when it relies on a single ASR round-trip word error rate (WER). A

PaTAS: A Framework for Trust Propagation in Neural Networks Using Subjective Logic

Model ReleasesDGX agent

arXiv:2511.20586v4 Announce Type: replace Abstract: Trustworthiness has become a key requirement for the deployment of artificial intelligence systems in safety-critical applications. Conventional eva

PATE-TabTransGAN: Differentially Private Synthetic Tabular Data Generation via Transformer-Based Student Discrimination

ResearchDGX agent

arXiv:2605.26802v1 Announce Type: new Abstract: Generating high-fidelity synthetic tabular data under formal differential privacy guarantees remains an open challenge. Methods that provide strong theo

Periodic Topological Deep Learning for Polymer Design and Discovery

Model ReleasesDGX agent

arXiv:2605.26833v1 Announce Type: cross Abstract: Polymers underpin applications across energy, healthcare, and materials science, yet their vast chemical space makes systematic discovery challenging.

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark

Model ReleasesDGX agent

arXiv:2506.00250v4 Announce Type: replace Abstract: Large Language Models (LLMs) have achieved remarkable performance on a wide range of Natural Language Processing (NLP) benchmarks, often surpassing

Persistent AI Agents in Academic Research: A Single-Investigator Implementation Case Study

Local AiDGX agent

arXiv:2605.26870v1 Announce Type: cross Abstract: Background: Large language models are typically evaluated as models, benchmarks, or short conversational episodes. Less is known about what happens wh

PersLitEval: Fine-grained Benchmark and Evaluation of LLMs on Persian Literature Questions

Model ReleasesDGX agent

arXiv:2605.27015v1 Announce Type: new Abstract: Despite impressive multilingual capabilities, large language models (LLMs) remain poorly evaluated on literary knowledge in non-English languages. We in

Persona Generators: Generating Diverse Synthetic Personas for Arbitrary Contexts

AgentsDGX agent

arXiv:2602.03545v2 Announce Type: replace Abstract: Evaluating AI systems that interact with humans requires understanding their behavior across diverse user populations, but collecting representative

Persona2Web: Benchmarking Personalized Web Agents for Contextual Reasoning with User History

Model ReleasesDGX agent

arXiv:2602.17003v2 Announce Type: replace-cross Abstract: Large language models have advanced web agents, yet current agents lack personalization capabilities. Since users rarely specify every detail

Personalized Generative Models for Contextual Debiasing

TutorialsDGX agent

arXiv:2605.26353v1 Announce Type: cross Abstract: Different visual patterns appear with different frequencies in the world: e.g., beach balls appear on sand more often than they do on a road. These st

Personalizing Embodied Multimodal Large Language Model Agents over Long-term User Interactions

AgentsDGX agent

arXiv:2605.26256v1 Announce Type: new Abstract: Multimodal large language model (MLLM)-based embodied agents have shown strong potential for solving complex tasks in physical environments. However, pe

Phase-Type Variational Autoencoders for Heavy-Tailed Data

ApplicationsDGX agent

arXiv:2603.01800v2 Announce Type: replace-cross Abstract: Heavy-tailed distributions are ubiquitous in real-world data, where rare but extreme events dominate risk and variability. However, standard V

PhyGHT: Physics-Guided HyperGraph Transformer for Signal Purification at the HL-LHC

Local AiDGX agent

arXiv:2602.20475v2 Announce Type: replace-cross Abstract: The High-Luminosity Large Hadron Collider (HL-LHC) at CERN will produce unprecedented datasets capable of revealing fundamental properties of

PhyPush: One Push is All You Need for Sensorless Physical Property Estimation with Physics-Guided Transformers

ApplicationsDGX agent

arXiv:2605.26284v1 Announce Type: new Abstract: Accurately estimating object mass and friction is fundamental to achieving reliable and adaptive robotic manipulation. Although interactive perception p

'PhyWorldBench': A Comprehensive Evaluation of Physical Realism in Text-to-Video Models

Model ReleasesDGX agent

arXiv:2507.13428v3 Announce Type: replace-cross Abstract: Video generation models have achieved remarkable progress in creating high-quality, photorealistic content. However, their ability to accurate

PICACO: Pluralistic In-Context Value Alignment of LLMs via Total Correlation Optimization

SafetyDGX agent

arXiv:2507.16679v3 Announce Type: replace-cross Abstract: In-Context Learning has shown great potential for aligning Large Language Models (LLMs) with human values, helping reduce harmful outputs and

PIDM-DP: Physics-Informed Diffusion with Dormand-Prince Integration for Chaotic System Identification and State Reconstruction across Multiple Dynamical Regimes

Model ReleasesDGX agent

arXiv:2605.26619v1 Announce Type: new Abstract: Reconstructing continuous state trajectories of chaotic dynamical systems from sparse, noisy observations remains a fundamental open problem in nonlinea

PILOT: A Data-Free Continual Learning Approach for Real-Time Semantic Segmentation via Boundary Guidance

TutorialsDGX agent

arXiv:2605.27128v1 Announce Type: new Abstract: Real-time semantic segmentation models offer an excellent balance between accuracy and inference speed. However, deploying these models in dynamic real

PilotTTS: A Disciplined Modular Recipe for Competitive Speech Synthesis

Model ReleasesDGX agent

arXiv:2605.27258v1 Announce Type: cross Abstract: Building state-of-the-art text-to-speech (TTS) systems typically demands millions of hours of proprietary data and complex multi-stage architectures,

PinPoint: Prompting with Informative Interior Points

ResearchDGX agent

arXiv:2605.26689v1 Announce Type: cross Abstract: Modern referring image segmentation pipelines couple a vision-language model (VLM) for grounding with a promptable segmenter such as the Segment Anyth

PitchBench: Measuring Pitch Hearing in Audio-Language Models

Model ReleasesDGX agent

arXiv:2605.26176v1 Announce Type: cross Abstract: Audio-language models (ALMs) are increasingly used in real-world applications that require understanding music, from music tutoring and transcription

PLAID: A Unified Data Model for Machine Learning on Heterogeneous Physics Simulations

ResearchDGX agent

arXiv:2505.02974v3 Announce Type: replace Abstract: Machine learning-based surrogate models have emerged as a powerful tool to accelerate simulation-driven scientific workflows, but their adoption is

Plan Then Action:High-Level Planning Guidance Reinforcement Learning for LLM Reasoning

Local AiDGX agent

arXiv:2510.01833v2 Announce Type: replace Abstract: Large language models (LLMs) demonstrate strong reasoning abilities via Chain-of-Thought (CoT), but their token-level generation encourages local de

Planning Neural Dynamics with Lie Group Embedding through Supervised Projective Manifold Learning

TutorialsDGX agent

arXiv:2605.26167v1 Announce Type: cross Abstract: We propose Lie group embedded dynamical neural networks (LieEDNN) and the corresponding learning algorithms based on gradient descent and metric proje

Plans for Evaluating Structured Generative Search Summaries

ResearchDGX agent

arXiv:2605.26400v1 Announce Type: cross Abstract: We propose a framework for evaluating structured generative search summaries that are placed atop organic web search results. A structured summary, ge

PlayClass: Automated Play Behaviour Classification in Poultry

ResearchDGX agent

arXiv:2605.27304v1 Announce Type: new Abstract: Automated monitoring of animal welfare has largely targeted negative indicators, leaving positive welfare behaviours such as play underexplored. To addr

PolyFusionAgent: A Multimodal Foundation Model and Autonomous AI Assistant for Polymer Property Prediction and Inverse Design

AgentsDGX agent

arXiv:2605.26543v1 Announce Type: new Abstract: Polymer discovery is central to fields ranging from energy storage to biomedicine, but it is hindered by an astronomically large chemical design space a

Pop-Up Distractions Reveal Bag-of-Events Behavior in Video Large Language Models

ResearchDGX agent

arXiv:2605.27101v1 Announce Type: cross Abstract: A key capability for video understanding is reliably linking subjects to events across time, yet whether Video Large Language Models (VideoLLMs) actua

Position: AI Safety Requires Effective Controllability

Model ReleasesDGX agent

arXiv:2605.27117v1 Announce Type: new Abstract: AI safety is still largely framed as alignment: training models to follow human preferences, safety policies, and normative constraints. That framing ha

Position: Machine Learning for Heart Transplant Allocation Policy Optimization Should Account for Incentives

SafetyDGX agent

arXiv:2602.04990v3 Announce Type: replace Abstract: The allocation of scarce donor organs constitutes one of the most consequential algorithmic challenges in healthcare. While the field is rapidly tra

Practical Anonymous Two-Party Gradient Boosting Decision Tree

SafetyDGX agent

arXiv:2605.26903v1 Announce Type: cross Abstract: Structured data is well handled by gradient-boosted decision trees (GBDT), which are usually trained on vertically partitioned features across mutuall

PRBench: A Standardized Probabilistic Robustness Benchmark

Model ReleasesDGX agent

arXiv:2511.01724v3 Announce Type: replace Abstract: Deep learning models are notoriously vulnerable to imperceptible perturbations. Most existing research centers on adversarial robustness (AR), which

Pretrained Approximators for Low-Thrust Trajectory Cost and Reachability

Model ReleasesDGX agent

arXiv:2605.26790v1 Announce Type: new Abstract: Low-thrust trajectory design relies heavily on repeated evaluations of fuel consumption and transfer feasibility, which require expensive optimal contro

Pretraining Data Exposure in Large Language Models: A Survey of Membership Inference, Data Contamination, and Security Implications

ResearchDGX agent

arXiv:2605.26133v1 Announce Type: cross Abstract: Large Language Models (LLMs) have become the predominant paradigm in NLP, advancing both research and industry. As model sizes and pretraining data gr

PRISM: A Multi-Dimensional Benchmark for Evaluating LLM Peer Reviewers

Model ReleasesDGX agent

arXiv:2605.26730v1 Announce Type: new Abstract: The rapid growth in submissions to machine learning venues has strained the scientific peer-review system and intensified interest in LLM-based automate

PRISM: Position-encoded Regressive Inverse Spectral Model for Multilayer Thin-Film Design

Model ReleasesDGX agent

arXiv:2605.26502v1 Announce Type: new Abstract: The inverse problem of multilayer thin-film optical coatings design represents a complex combinatorial-continuous optimization challenge. We present PRI

Probabilistic Recurrent Intention Switching Model

ResearchDGX agent

arXiv:2605.26998v1 Announce Type: new Abstract: Inverse reinforcement learning (IRL) recovers reward functions from observed behavior, yet traditional methods assume a single stationary reward that ca

Probabilistic Smoothing with Ratio-Monotone Transforms for Global Optimization

ResearchDGX agent

arXiv:2605.27316v1 Announce Type: new Abstract: Probabilistic smoothing is a standard tool for global optimization, but existing methods rely on Gaussian kernels and specific transforms, often resulti

Probing Cultural Awareness in LLMs: A Case Study of Cross-Culture Aesthetic Stylistics

Model ReleasesDGX agent

arXiv:2605.27296v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed in diverse cultural contexts, yet their ability to master aesthetic stylistics, i.e., the strateg

Probing Minimalist Phase Structure in LLMs: What Universal Dependencies Cannot Represent

ResearchDGX agent

arXiv:2605.26431v1 Announce Type: new Abstract: Structural probes train on Universal Dependencies (UD), which does not encode formal-syntactic abstractions such as phase boundaries or phase-internal c

Probing the Knowledge Boundary: An Interactive Agentic Framework for Deep Knowledge Extraction

AgentsDGX agent

arXiv:2602.00959v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) can be seen as compressed knowledge bases, but it remains unclear what knowledge they truly contain and how far t

Prompt Injection Detection is Regime-Dependent: A Deployment-Aware Evaluation with Interpretable Structural Signals

Model ReleasesDGX agent

arXiv:2605.26999v1 Announce Type: new Abstract: Prompt injection poses a critical threat to the safe deployment of large language models, yet existing detection approaches are typically evaluated unde

← Previous
1…583584585586587…1040
Next →