AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,037 results
28 May 2026

Locality-Aware Redundancy Pruning for LLM Depth Compression

Local AiDGX agent

arXiv:2605.27786v1 Announce Type: cross Abstract: Large language models are known to contain representational redundancy across network depth, making depth pruning an effective approach for improving

Long Live The Balance: Information Bottleneck Driven Tree-based Policy Optimization

SafetyDGX agent

arXiv:2605.28109v1 Announce Type: new Abstract: Recent advances in online reinforcement learning (RL) for large language models (LLMs) have demonstrated promising performance in complex reasoning task

Machine Learning methods for event classification and vertex reconstruction of the 12C + 12C reaction with the MATE-TPC

ResearchDGX agent

arXiv:2605.28296v1 Announce Type: new Abstract: In modern nuclear physics experiments, identifying events of interest is challenging for nuclear reaction studies with the active target Time Projection

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

MGRetrieval: Memory-Guided Reflective Retrieval for Long-Term Dialogue Agents

ResearchDGX agent

arXiv:2605.27437v1 Announce Type: cross Abstract: Large Language Models (LLMs) have made significant progress in dialogue, yet redundant memory contexts severely limit their effectiveness in long-term

Misalignment Between Backpropagation and the Hierarchy of Brain Responses to Images

TutorialsDGX agent

arXiv:2605.28693v1 Announce Type: cross Abstract: Backpropagation is the core learning mechanism underlying deep learning. However, whether and how this algorithm is implemented in the brain remains h

Mobile-Aptus: Confidence-Driven Proactive and Robust Interaction in MLLM-based Mobile-Using Agents

SafetyDGX agent

arXiv:2605.28629v1 Announce Type: new Abstract: Recent advancements in multimodal large language models (MLLMs) have shown exceptional potential in enabling mobile-using agents to autonomously execute

Neural Quantum Spectral Operator Learning for Solving Partial Differential Equations

ResearchDGX agent

arXiv:2605.27408v1 Announce Type: cross Abstract: Partial differential equations (PDEs) are central to modeling physical and engineering systems, but repeatedly solving parametric PDEs remains computa

📢 New @heyjasper release ! 📢 MONET 🌸 : An Apache2.0 deduped and recaptioned dataset of 105M samples unlocking reproducible text-to-image …

IndustryDGX agent

📢 New @heyjasper release ! 📢 MONET 🌸 : An Apache2.0 deduped and recaptioned dataset of 105M samples unlocking reproducible text-to-image research. Nano T2I 🖌️ : A codebase to train your own T2I model

Off-Policy Learning to Reason Works Because It Is More Pessimistic Than You Think

SafetyDGX agent

arXiv:2605.28150v1 Announce Type: new Abstract: Large scale reinforcement learning has become a central tool for improving reasoning in large language models. At this scale, generation is often lagged

OmniEgo-R^2: A Routed Reasoning Framework for the 1st Cross-Domain EgoCross Challenge at CVPR 2026

ResearchDGX agent

arXiv:2605.24481v2 Announce Type: replace Abstract: The 1st Cross-Domain EgoCross Challenge at EgoVis, CVPR 2026 evaluates whether multimodal large language models can reason over egocentric videos ac

On the Learnability of Test-Time Adaptation: A Recovery Complexity Perspective

ResearchDGX agent

arXiv:2605.28057v1 Announce Type: cross Abstract: Test-time adaptation (TTA) aims to adapt models to maintain reliable performance on non-stationary test streams without requiring labeled data. Despit

Optimal and Diffusion Transports in Machine Learning

ResearchDGX agent

arXiv:2512.06797v2 Announce Type: replace-cross Abstract: Several problems in machine learning are naturally expressed as the design and analysis of time-evolving probability distributions. This inclu

Opus 4.8 on AI Gateway

ToolsDGX agent

Vercel announced support for Claude Opus 4.8 on its AI Gateway, enabling developers to integrate Anthropic's latest large language model through Vercel's unified API platform. This update allows seaml

Orchid Security targets AI agent sprawl with new identity governance tools

AgentsDGX agent

Orchid Security Inc. today extended its Identity Control Plane with a set of capabilities aimed at governing artificial intelligence agents, saying existing identity and access management models canno

Perplexity Computer is now available inside Microsoft Excel, Word, PowerPoint, and Outlook. Orchestrate across work with Computer directly i…

ToolsDGX agent

Perplexity Computer is now available inside Microsoft Excel, Word, PowerPoint, and Outlook. Orchestrate across work with Computer directly in the side panel of your app to draft documents, model, buil

Picid: A Modular Evaluation Infrastructure for Reproducible PHM Across Tasks and Domains

SafetyDGX agent

arXiv:2605.28345v1 Announce Type: new Abstract: Progress in Prognostics and Health Management (PHM) is hindered by the lack of standardized and reusable evaluation practices across tasks, datasets, an

Quality-constrained Entropy Maximization Policy Optimization for LLM Diversity

SafetyDGX agent

arXiv:2602.15894v2 Announce Type: replace Abstract: In many large language model (LLM) alignment applications, users expect not only high-quality outputs but also substantial diversity. However, exist

Revealing Algorithmic Deductive Circuits for Logical Reasoning

TutorialsDGX agent

arXiv:2605.27824v1 Announce Type: new Abstract: Recent studies have shown that Large Language Models (LLMs) can achieve strong reasoning performance by incorporating functional symbolic representation

Securing Retrieval-Augmented Generation: A Taxonomy of Attacks, Defenses, and Future Directions

AgentsDGX agent

arXiv:2604.08304v2 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) extends large language models (LLMs) with external knowledge, but this access path also introduces securi

Semantic-Aware Interpretable Multimodal Music Auto-Tagging

ResearchDGX agent

arXiv:2505.17233v3 Announce Type: replace Abstract: Music auto-tagging is essential for organizing and discovering music in extensive digital libraries. While foundation models achieve exceptional per

Semantic Flow Regularization: Teaching LLMs to Generate Diverse Yet Coherent Responses

ResearchDGX agent

arXiv:2605.27971v1 Announce Type: cross Abstract: When large language models are fine-tuned to generate persona- or tone-conditioned responses, their output diversity is severely limited--a failure we

Semantic Optimal Transport for Sparse Autoencoder Feature Matching and Circuit Compression

ResearchDGX agent

arXiv:2605.28567v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) have become a central tool for interpreting language models. However, two key SAE analyses that remain difficult to scale a

Skill0.5: Joint Skill Internalization and Utilization for Out-of-Distribution Generalization in Agentic Reinforcement Learning

AgentsDGX agent

arXiv:2605.28424v1 Announce Type: new Abstract: Equipping large language models with explicit skills has emerged as a promising paradigm for enabling autonomous agents to solve complex tasks. Agent sk

slide from

AgentsDGX agent

slide from The Redpoint InfraRed 100 is now live. These are the companies building the infrastructure that powers everything happening in AI right now, from world models and agent runtimes to the sand

Smoothed Score Queries and the Complexity of Sampling

ResearchDGX agent

arXiv:2605.27769v1 Announce Type: cross Abstract: We study the query complexity of sampling from high-dimensional Gaussian distributions using gradient information. In the standard oracle model, exact

Soft-SVeRL: Self-Verified Reinforcement Learning with Soft Rewards

SafetyDGX agent

arXiv:2605.28561v1 Announce Type: new Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) has improved language models in domains such as mathematics and code, where correctness can be che

SPARD: Defending Harmful Fine-Tuning Attack via Safety Projection with Relevance-Diversity Data Selection

SafetyDGX agent

arXiv:2605.28030v1 Announce Type: cross Abstract: Fine-tuning large language models often undermines their safety alignment, a problem further amplified by harmful fine-tuning attacks in which adversa

STFlow: Data-Coupled Flow Matching for Geometric Trajectory Simulation

ResearchDGX agent

arXiv:2505.18647v3 Announce Type: replace-cross Abstract: Simulating trajectories of dynamical systems is a fundamental problem in a wide range of fields such as molecular dynamics, biochemistry, and

Super-Resolved Canopy Height Mapping from Sentinel-2 Time Series Using Airborne LiDAR HD Reference Data across Metropolitan France

ResearchDGX agent

arXiv:2512.11524v3 Announce Type: replace Abstract: Fine-scale forest monitoring is essential for understanding canopy structure and its dynamics, which are key indicators of carbon stocks, biodiversi

The Shape of Overthinking: Backtracking Bursts in Long Reasoning Traces

Local AiDGX agent

arXiv:2605.27965v1 Announce Type: new Abstract: Reasoning models often generate long traces in which useful self-correction and unproductive revision are hard to distinguish. We study this distinction

Understanding Automated Program Repair Agents Through the Lens of Traceability: An Empirical Study

AgentsDGX agent

arXiv:2506.08311v2 Announce Type: replace-cross Abstract: Automated Program Repair (APR) agents leverage Large Language Models (LLMs) to autonomously diagnose and fix software bugs through reasoning,

Uni-LaViRA: Language-Vision-Robot Actions Translation for Unified Embodied Navigation

HardwareDGX agent

arXiv:2605.27582v1 Announce Type: cross Abstract: Embodied navigation requires an agent to map language and visual observations to a stream of spatial actions that drive a real robot through environme

Utility-Aware Multimodal Contrastive Learning for Product Image Generation

SafetyDGX agent

arXiv:2605.28733v1 Announce Type: new Abstract: Product images strongly influence consumer decision-making in online marketplaces. Empowered by multimodal contrastive learning, generative AI can outpu

ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention

ResearchDGX agent

arXiv:2602.07574v2 Announce Type: replace-cross Abstract: Modern multimodal large language models (MLLMs) adopt a unified self-attention design that processes visual and textual tokens at every Transf

When NPUs Are Not Always Faster: A Stage-Level Analysis of Mobile LLM Inference

Local AiDGX agent

arXiv:2605.27435v1 Announce Type: cross Abstract: Deploying large language models (LLMs) on mobile devices increasingly relies on heterogeneous execution, yet no prior study has systematically charact

When Seekers Are Hard to Help: Evaluating Emotional Support Dialogue Systems in Worst-Case Interactions

ResearchDGX agent

arXiv:2605.28228v1 Announce Type: new Abstract: Emotional Support Dialogue Systems (ESDSes) are increasingly evaluated and trained with LLM-simulated seekers. However, such simulated seekers often beh

xKV: Cross-Layer KV-Cache Compression via Aligned Singular Vector Extraction

SafetyDGX agent

arXiv:2503.18893v2 Announce Type: replace Abstract: Long-context Large Language Models (LLMs) enable powerful applications but incur high memory costs due to the key-value states (KV-Cache). Recent st

27 May 2026

9/ Real-time RL is where it gets fun. Catch live signals from real users on real generations. Update continuously. Ship a new version every …

ToolsDGX agent

9/ Real-time RL is where it gets fun. Catch live signals from real users on real generations. Update continuously. Ship a new version every few hours. Only works if the base model is already good enou

A Physics-Informed Hierarchical Neural Network for Microwave Scattering Analysis of 3D PEC Targets

ResearchDGX agent

arXiv:2508.03774v5 Announce Type: replace-cross Abstract: Accurate modeling of scattering from three-dimensional (3D) perfectly electrically conducting (PEC) targets at microwave frequencies constitut

Accountable Human-AI Deliberation with LLMs: Scaling Collective Intelligence through Symbiotic Scaffolding

ResearchDGX agent

arXiv:2605.26940v1 Announce Type: new Abstract: Large language models (LLMs) can support democratic deliberation at scales previously constrained by turn-taking and facilitation bandwidth. Recent work

AD-H: Language-guided Autonomous Driving with Hierarchical Agents

AgentsDGX agent

arXiv:2406.03474v2 Announce Type: replace Abstract: Language-guided autonomous driving requires bridging a large abstraction gap between high-level natural-language instructions and low-level vehicle

Adaptive Multi-prompt Contrastive Network for Few-shot Out-of-distribution Detection

ApplicationsDGX agent

arXiv:2506.17633v2 Announce Type: replace-cross Abstract: Out-of-distribution (OOD) detection attempts to distinguish outlier samples to prevent models trained on the in-distribution (ID) dataset from

Advancing Metallic Surface Defect Detection via Anomaly-Guided Pretraining on a Large Industrial Dataset

ResearchDGX agent

arXiv:2509.18919v2 Announce Type: replace Abstract: The pretraining-finetuning paradigm is a crucial strategy in metallic surface defect detection for mitigating the challenges posed by data scarcity.

Approximate Equivariance via Projection-based Regularisation

SafetyDGX agent

arXiv:2601.05028v2 Announce Type: replace Abstract: Equivariance is a powerful inductive bias in neural networks, improving generalisation and physical consistency. Recently, however, non-equivariant

Assessing Per-Sample Membership Inference Vulnerability without Retraining

ResearchDGX agent

arXiv:2602.15919v2 Announce Type: replace-cross Abstract: Recent work in the privacy literature shows that sample-targeted membership inference attacks (MIAs) significantly outperform untargeted appro

ATOM: Instantiating Budget-Controllable Multi-Agent Collaboration via Nucleus-Electron Hierarchy

AgentsDGX agent

arXiv:2605.26178v1 Announce Type: cross Abstract: Large Language Model (LLM)-based multi-agent systems rely on optimized collaboration topologies to balance performance and communication costs. Howeve

BASIS: Batchwise Advantage Estimation from Single-Rollout Information Sharing for LLM Reasoning

SafetyDGX agent

arXiv:2605.27293v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards has become a standard recipe for improving the reasoning abilities of large language models. Existing alg

Beyond Self-Talk: A Communication-Centric Survey of LLM-Based Multi-Agent Systems

AgentsDGX agent

arXiv:2502.14321v3 Announce Type: replace-cross Abstract: Large language model-based multi-agent systems have recently gained significant attention due to their potential for complex, collaborative, a

Beyond Trajectory-Level Attribution: Graph-Based Credit Assignment for Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2605.26684v1 Announce Type: cross Abstract: Group-based reinforcement learning (RL) methods have achieved remarkable success in improving the performance of large language models (LLMs) and have

Bilevel Optimization over Saddle Points of Zero-Sum Markov Games

SafetyDGX agent

arXiv:2605.26654v1 Announce Type: cross Abstract: Reinforcement learning (RL) often has a hierarchical structure, where an upper-level (UL) learner selects model parameters and a lower-level (LL) deci

Bounded Path Context: A Controlled Study of Visible Path History in LLM-Based Knowledge Graph Question Answering

Local AiDGX agent

arXiv:2605.26645v1 Announce Type: new Abstract: LLM-based knowledge-graph question answering (KGQA) delegates graph traversal to language models, turning each question into a sequence of local relatio

Cisco and OpenAI redefine enterprise engineering with Codex

ApplicationsDGX agent

Cisco and OpenAI partnered to integrate OpenAI's Codex AI model into Cisco's enterprise engineering tools to enhance software development and automation capabilities. The collaboration aims to improve

Convergence of Spectral Descent for Non-smooth Optimization

ResearchDGX agent

arXiv:2605.26977v1 Announce Type: new Abstract: The Muon optimizer has recently demonstrated remarkable empirical success in training large language models. However, the theoretical understanding of i

CRoFT: Robust Fine-Tuning with Concurrent Optimization for OOD Generalization and Open-Set OOD Detection

TutorialsDGX agent

arXiv:2405.16417v2 Announce Type: replace Abstract: Recent vision-language pre-trained models (VL-PTMs) have shown remarkable success in open-vocabulary tasks. However, downstream use cases often invo

CSV-ViT: A Vision Transformer with the Variable-sized Cortical Supervertices for Detection of Alzheimer's Disease Pathologies

ResearchDGX agent

arXiv:2605.26514v1 Announce Type: cross Abstract: Confirming Alzheimer's disease (AD) typically relies on positron emission tomography (PET), which remains costly and invasive, motivating the use of s

Data-driven sparse identification of governing PDEs via knockoff filters and multi-criteria trade-offs

ResearchDGX agent

arXiv:2605.26631v1 Announce Type: cross Abstract: We propose KO-PDE-IDENT, a data-driven framework for identifying parsimonious partial differential equations (PDEs) with false discovery rate (FDR) co

Deep Learning-based Algebraic Reynolds Stress Closures for RANS Simulations of Turbulent Flows

ResearchDGX agent

arXiv:2605.26358v1 Announce Type: cross Abstract: Turbulence is ubiquitous in engineering and science, yet direct simulation is prohibitively expensive. The Reynolds-averaged Navier-Stokes (RANS) equa

Detached Skip-Links and R-Probe: Decoupling Feature Aggregation from Gradient Propagation for MLLM OCR

ResearchDGX agent

arXiv:2603.20020v2 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) excel at high-level reasoning yet fail on OCR tasks where fine-grained visual details are compromised

Detectability in Diversity: Improved Canary Crafting for Privacy Auditing in One Run

ResearchDGX agent

arXiv:2605.27292v1 Announce Type: new Abstract: Privacy auditing aims to empirically assess privacy leakage in machine learning models using membership inference attacks (MIAs), and to derive lower bo

Detecting Is Not Resolving: The Monitoring Control Gap in Retrieval Augmented LLMs

SafetyDGX agent

arXiv:2605.27157v1 Announce Type: new Abstract: Retrieval-augmented LLMs are deployed for tasks where evidence quality determines action safety, yet evaluation protocols assume that single-turn robust

← Previous
1…762763764765766…1018
Next →