AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
Human
87,814Total entries
1Added by human
87,813Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,480 results
28 May 2026

Optimal LTLf Synthesis

Model ReleasesDGX agent

arXiv:2605.11544v2 Announce Type: replace Abstract: Strategy synthesis typically follows an all-or-nothing paradigm, returning unrealisable whenever a specification cannot be guaranteed in an uncertai

Optimal ridge regularization revisited

Model ReleasesDGX agent

arXiv:2605.28679v1 Announce Type: new Abstract: We consider L^2-regularized linear (ridge) regression over a finite data sample X with bounded covariance and linear prediction targets y with additive

OR-Space: A Full-Lifecycle Workspace Benchmark for Industrial Optimization Agents

Model ReleasesDGX agent

arXiv:2605.28158v1 Announce Type: new Abstract: Large language model (LLM) agents are increasingly used to assist with operations research (OR) modeling, yet existing OR-oriented benchmarks often redu

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

OralAgent: Integrating Reasoning, Tools, and Knowledge for Interactive Dental Image Analysis

Model ReleasesDGX agent

arXiv:2605.27378v1 Announce Type: new Abstract: Dental image analysis plays a pivotal role in supporting accurate diagnosis and treatment planning in oral healthcare. Although recent advances have pro

OSP-Next: Efficient High-Quality Video Generation with Sparse Sequence Parallelism, HiF8 Quantization, and Reinforcement Learning

Local AiDGX agent

arXiv:2605.28691v1 Announce Type: new Abstract: Diffusion Transformers achieve strong video generation quality, but the quadratic cost of full attention limits efficiency. We introduce OSP-Next, an ef

Out of Sight, Not Out of Mind: Unveiling Latent Attack in Latent-based Multi-Agent Systems

AgentsDGX agent

arXiv:2605.28214v1 Announce Type: cross Abstract: Latent-based multi-agent systems replace parts of explicit inter-agent communication with hidden representations, offering a new direction for efficie

Outer-Momentum Restarting in High-Dimensional Two-Phase Optimization

Local AiDGX agent

arXiv:2605.28585v1 Announce Type: new Abstract: Communication-efficient distributed optimizers such as DiLoCo reduce synchronization costs by letting workers perform many local updates before aggregat

Parameter-Efficient Generative Modeling with Controlled Vector Fields

Model ReleasesDGX agent

arXiv:2605.28267v1 Announce Type: new Abstract: We introduce a continuous-time generative modeling framework, motivated by the Chow-Rashevskii theorem, that builds expressive flows from a small set of

Paraphrase Brittleness in Production Retrieval-Augmented Commercial Recommendation: Reproducibility Below the Rerun-Stability Baseline

Model ReleasesDGX agent

arXiv:2605.27440v1 Announce Type: cross Abstract: Small changes to how a buyer phrases a question -- 'best CRM' vs 'top CRM' vs 'best CRM for a SaaS startup' -- produce substantially different brand r

Particle-Guided Diffusion Models for Partial Differential Equations

Model ReleasesDGX agent

arXiv:2601.23262v2 Announce Type: replace Abstract: We introduce a guided stochastic sampling method that augments sampling from diffusion models with physics-based guidance derived from partial diffe

PAST2HARM: A Simple Adaptive Past Tense Attack for Jailbreaking Multimodal AI

Model ReleasesDGX agent

arXiv:2605.27545v1 Announce Type: new Abstract: Jailbreak attacks on multimodal AI systems remain underexplored, even though unsafe image generation can have more severe consequences than unsafe text

Patched-DeltaNet: Token-Level Event-Driven Memory for Linear-Time Anomaly Detection

Model ReleasesDGX agent

arXiv:2605.27992v1 Announce Type: new Abstract: Time series anomaly detection is critical for maintaining the reliability of mission-critical systems. While Transformer-based models like PatchTST have

Path Channels and Plan Extension Kernels: a Mechanistic Description of Planning in a Sokoban RNN

ResearchDGX agent

arXiv:2506.10138v3 Announce Type: replace-cross Abstract: We partially reverse-engineer a convolutional recurrent neural network (RNN) trained with model-free reinforcement learning to play the box-pu

Pattern Recognition Tasks with Personalized Federated Learning

ResearchDGX agent

arXiv:2605.27816v1 Announce Type: new Abstract: Personalized Federated Learning (PFL) constitutes a novel paradigm that tailors Machine Learning (ML) models to individual clients, thereby furnishing p

PEAM: Parametric Embodied Agent Memory through Contrastive Internalization of Experience in Minecraft

Model ReleasesDGX agent

arXiv:2605.27762v1 Announce Type: new Abstract: We present PEAM, a Parametric Embodied Agent Memory framework in Minecraft that transforms agent memory from inference-time retrieval into parameter-res

PEAR: Equal Area Weather Forecasting on the Sphere

ResearchDGX agent

arXiv:2505.17720v3 Announce Type: replace Abstract: Artificial intelligence is rapidly reshaping the natural sciences, with weather forecasting emerging as a flagship AI4Science application where mach

PEAR: Pairwise Evaluation for Automatic Relative Scoring in Machine Translation

Model ReleasesDGX agent

arXiv:2601.18006v2 Announce Type: replace Abstract: We present PEAR (Pairwise Evaluation for Automatic Relative Scoring), a supervised quality estimation (QE) metric family that reframes reference-fre

PEFT-Arena: Understanding Parameter-Efficient Finetuning from a Stability-Plasticity Perspective

Model ReleasesDGX agent

arXiv:2605.28819v1 Announce Type: cross Abstract: Parameter-efficient finetuning (PEFT) has become the standard approach for adapting large language models, yet evaluations largely emphasize downstrea

Performance and Explainability Requirements of Evolutionary Algorithms in Real-World Physics-Informed Optimization

ApplicationsDGX agent

arXiv:2605.28164v1 Announce Type: cross Abstract: Evolutionary computation offers a variety of tools to solve complex real-world optimization problems. However, research often focuses on smaller, simp

Periodic RoPE for Infinite Context LLMs

Model ReleasesDGX agent

arXiv:2605.27980v1 Announce Type: cross Abstract: The ability to process ultra-long contexts is crucial for large language models (LLMs) to perform long-horizon tasks. While recent efforts have extend

Personal Visual Memory from Explicit and Implicit Evidence

Model ReleasesDGX agent

arXiv:2605.28806v1 Announce Type: cross Abstract: Long-term memory is increasingly important for personalized AI agents, yet existing benchmarks and methods remain largely text-centric. Even when imag

Personality, Role, and Expressive Style in Large Language Models: An Interactionist Analysis

AgentsDGX agent

arXiv:2605.28037v1 Announce Type: new Abstract: Prompt-based personality control is a key technique for designing large language model (LLM) dialogue agents that behave consistently across social cont

Personalized Observation Normalization for Federated Reinforcement Learning in Simulation Environments with Heterogeneity

Model ReleasesDGX agent

arXiv:2605.27385v1 Announce Type: cross Abstract: Federated reinforcement learning (FedRL) enables multiple agents to collaboratively train a global policy without sharing raw data, making it ideal fo

Persuade Me if You Can: A Framework for Evaluating Persuasion Effectiveness and Susceptibility Among Large Language Models

Model ReleasesDGX agent

arXiv:2503.01829v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) demonstrate persuasive capabilities that rival human-level persuasion. While these capabilities can be used for s

PetroBench: A Benchmark for Large Language Models in Petroleum Engineering

Model ReleasesDGX agent

arXiv:2605.28032v1 Announce Type: new Abstract: Large Language Models are increasingly applied in the petroleum industry, highlighting the need for a domain-specific evaluation framework. This study d

PhAME: Phenotype-Aware Molecular Editing via Latent Diffusion

ResearchDGX agent

arXiv:2605.28226v1 Announce Type: new Abstract: Small-molecule drug discovery requires simultaneous optimization of numerous properties of candidate molecules. These properties can be investigated thr

Picid: A Modular Evaluation Infrastructure for Reproducible PHM Across Tasks and Domains

SafetyDGX agent

arXiv:2605.28345v1 Announce Type: new Abstract: Progress in Prognostics and Health Management (PHM) is hindered by the lack of standardized and reusable evaluation practices across tasks, datasets, an

PINE: Pruning Boosted Tree Ensembles with Conformal In-Distribution Prediction Equivalence

Model ReleasesDGX agent

arXiv:2605.28068v1 Announce Type: new Abstract: Tree ensembles are machine learning models with strong predictive performance and interpretability, and remain widely used for tabular data. Standard pr

PIRS: Physics-Informed Reward Shaping for SAC-Based Building Energy Management

AgentsDGX agent

arXiv:2605.28232v1 Announce Type: new Abstract: Occupant comfort and grid-aware energy efficiency are competing objectives whose joint optimization depends critically on how reward functions are speci

Plan Before Search: Search Agents Need Plan

AgentsDGX agent

arXiv:2605.28354v1 Announce Type: new Abstract: Training large language models as retrieval-augmented reasoning agents typically combines reinforcement learning with an SFT cold start distilled from a

Planning a Community Approach to Diabetes Care in Low- and Middle-Income Countries Using Optimization

ResearchDGX agent

arXiv:2305.06426v2 Announce Type: replace Abstract: Diabetes is a global health priority, especially in low- and-middle-income countries, where over 50% of premature deaths are attributed to high bloo

Plant, Persist, Trigger: Sleeper Attack on Large Language Model Agents

Model ReleasesDGX agent

arXiv:2605.28201v1 Announce Type: new Abstract: Large Language Model (LLM) agents remain vulnerable to safety threats from the external environment, where attackers inject adversarial content into ext

Playing with Words, Improving with Rewards: Training Language Models for Creative Association

ResearchDGX agent

arXiv:2605.27832v1 Announce Type: new Abstract: Large Language Models (LLMs) are being applied to increasingly difficult problems and use cases. To navigate their vast solution spaces effectively, LLM

PLS in the Mirror of Self-Attention

ResearchDGX agent

arXiv:2605.28592v1 Announce Type: new Abstract: This note provides an interesting observation on casting partial least square (PLS) as a linearized self-attention so that PLS may be studied within the

Plug-and-Play Benchmarking of Reinforcement Learning Algorithms for Large-Scale Flow Control

Model ReleasesDGX agent

arXiv:2601.15015v2 Announce Type: replace Abstract: Reinforcement learning (RL) has shown promising results in active flow control (AFC), yet progress in the field remains difficult to assess as exist

PocketGS: On-Device Training of 3D Gaussian Splatting for High Perceptual Modeling

Local AiDGX agent

arXiv:2601.17354v5 Announce Type: replace Abstract: While 3D Gaussian Splatting (3DGS) enables real-time rendering, its training demands workstation-level compute and memory, making mobile deployment

POINav: Benchmarking and Enhancing Final-Meters Arrival in Real-World Vision-Language Navigation

Model ReleasesDGX agent

arXiv:2605.28237v1 Announce Type: cross Abstract: Real-world navigation is fundamentally driven by Points of Interest (POIs), yet reaching a precise POI remains a critical 'final-meters' challenge. Ex

PointQ-Bench: Benchmarking Diagnostic and Interpretable Point Cloud Quality Assessment

Model ReleasesDGX agent

arXiv:2605.28241v1 Announce Type: new Abstract: Point cloud quality plays a critical role in 3D acquisition, reconstruction, rendering, and perception, yet existing point cloud quality assessment (PCQ

Poison with Style: A Practical Poisoning Attack on Code Large Language Models

ResearchDGX agent

arXiv:2605.27631v1 Announce Type: cross Abstract: Code Large Language Models (CLLMs) serve as the core of modern code agents, enabling developers to automate complex software development tasks. In thi

PortBench: A Correlation-Aware, Full-Pipeline Benchmark for LLM-Driven Portfolio Management

Model ReleasesDGX agent

arXiv:2605.27887v1 Announce Type: new Abstract: LLMs have shown strong performance across diverse financial tasks, yet portfolio management (PM), a critical financial decision-making task, remains poo

Position: Retire the 'Positive Backdoor' Label -- Secret Alignment Requires Strict and Systematic Evaluation

SafetyDGX agent

arXiv:2605.28597v1 Announce Type: cross Abstract: This position paper argues that the AI/ML community should stop overclaiming and retire the label 'positive backdoor,' and instead treat trigger-activ

Preference-Shaped Expected Hypervolume and R2 Improvement: Exact Computation and Monotonicity

ResearchDGX agent

arXiv:2605.28746v1 Announce Type: cross Abstract: This paper studies preference-shaped expected improvement criteria for Bayesian multiobjective optimization. We consider two indicator families which

Prefix-Safe Bayesian Belief Tracking for LLM Reasoning Reliability:Separating Calibration from Ranking

ResearchDGX agent

arXiv:2605.27712v1 Announce Type: new Abstract: Long reasoning traces need reliability estimates before final answers are known. We study prefix-conditioned eventual-success estimation, P(y=1 mid o_{1

Pressure-Testing Deception Probes in LLMs: Scaling, Robustness, and the Geometry of Deceptive Representations

Model ReleasesDGX agent

arXiv:2605.27958v1 Announce Type: cross Abstract: Linear probes trained on LLM activations are increasingly proposed as deception-detection metrics, yet report AUROC exceeding 0.96 on clean benchmarks

PrimitiveVLA: Learning Reusable Motion Primitives for Efficient and Generalizable Robotic Manipulation

TutorialsDGX agent

arXiv:2605.28634v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models offer a promising paradigm for generalist robotic policies, yet their adaptation is hindered by data inefficiency an

Principled Algorithms for Optimizing Generalized Metrics in Multi-Label Learning

ApplicationsDGX agent

arXiv:2605.28767v1 Announce Type: new Abstract: Many real-world classification tasks require predicting multiple labels per instance, necessitating the optimization of complex evaluation metrics such

PrionNER: A Named Entity Recognition Dataset for Prion Disease Biomedical Literature

Model ReleasesDGX agent

arXiv:2605.28375v1 Announce Type: new Abstract: Prion diseases are rare, rapidly progressive, and fatal neurodegenerative disorders that remain difficult to diagnose, particularly in their early stage

Privacy Protection Against Personalized Text-to-Image Synthesis via Cross-image Consistency Constraints

ResearchDGX agent

arXiv:2504.12747v2 Announce Type: replace Abstract: The rapid advancement of diffusion models and personalization techniques has made it possible to recreate individual portraits from just a few publi

Privately Estimating Monotone Statistics in Polynomial Time

Model ReleasesDGX agent

arXiv:2605.27912v1 Announce Type: cross Abstract: We study efficient differentially private algorithms for estimating monotone statistics, i.e., statistics that are monotone under the addition of new

Probabilistic Data-Driven Modelling of Astrophysical Transients: The Neural Process Family for Ultrafast and Class-Agnostic Light Curve Reconstruction with NightLANP

Model ReleasesDGX agent

arXiv:2605.27527v1 Announce Type: cross Abstract: Astrophysical observations taken from Earth are subject to weather, environmental, and scientific constraints that lead to sparse, irregular light cur

Probability-Entropy Calibration: An Elastic Indicator for Adaptive Fine-tuning

SafetyDGX agent

arXiv:2602.01745v2 Announce Type: replace-cross Abstract: Token-level reweighting is a simple yet effective mechanism for controlling supervised fine-tuning, but common indicators are largely one-dime

Probing for Knowledge Attribution in Large Language Models

Model ReleasesDGX agent

arXiv:2602.22787v2 Announce Type: replace-cross Abstract: Large language model (LLM) hallucinations, meaning fluent but factually incorrect generations, fall into two types: faithfulness violations, w

Probing Social Identity Bias in Chinese LLMs with Gendered Pronouns and Social Groups

SafetyDGX agent

arXiv:2510.06974v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed in user-facing applications, raising concerns that they may reflect and amplify social biases

ProgVLA: Progress-Aware Robot Manipulation Skill Learning

Model ReleasesDGX agent

arXiv:2605.28231v1 Announce Type: cross Abstract: We present ProgVLA, a compact vision-language-action (VLA) model designed for reliable robot manipulation under tight compute and memory budgets. The

Prominence-Stratified Failure Modes in Retrieval-Augmented Commercial Recommendation: A 37,000-Run Audit

Model ReleasesDGX agent

arXiv:2605.27439v1 Announce Type: cross Abstract: AI assistants like ChatGPT and Claude are recommendation engines, not search engines: they answer commercial queries by directly nominating brands rat

Prompt Codebooks: Discrete Compositional Optimization for Language Model Instruction Refinement

Model ReleasesDGX agent

arXiv:2605.28360v1 Announce Type: new Abstract: Automatic prompt optimization (APO) has driven significant gains in LLM-based agentic workflows. However, existing methods treat each task's prompt as a

PromptEmbedder:: Efficient and Transferable Text Embedding via Dual-LLM Soft Prompting

Model ReleasesDGX agent

arXiv:2605.28066v1 Announce Type: cross Abstract: Large Language Models (LLMs) have demonstrated remarkable efficacy in text embedding, yet current adaptation methods like LoRA face significant bottle

Prompting Is All You Need: Multi-view Prompting Large Language Models for Aspect-Based Sentiment Analysis

Model ReleasesDGX agent

arXiv:2605.28058v1 Announce Type: new Abstract: Recent work explored the capabilities of Large Language Models (LLMs) in Aspect-Based Sentiment Analysis (ABSA) through few-shot prompting, requiring su

Proper Agnostic Learning of Functions of Halfspaces under Gaussian Marginals

ResearchDGX agent

arXiv:2605.27594v1 Announce Type: cross Abstract: We study the problem of computationally efficient proper agnostic learning of multidimensional concept classes under the Gaussian distribution. In thi

Proprio: Latent Self-Scoring and Inference-Time Refinement for Physically Plausible Video Generation

ResearchDGX agent

arXiv:2605.28230v1 Announce Type: new Abstract: Modern video generative models produce visually impressive results, yet frequently violate basic physical principles. We propose Proprio, a training-fre

← Previous
1…570571572573574…1042
Next →