AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
Human
88,271Total entries
1Added by human
88,270Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,897 results
29 May 2026

AgentDoG 1.5: A Lightweight and Scalable Alignment Framework for AI Agent Safety and Security

Model ReleasesDGX agent

arXiv:2605.29801v1 Announce Type: new Abstract: Modern open-world agents such as OpenClaw exhibit powerful cross-environment execution capabilities yet introduce broad new safety risk sources. Meanwhi

AgentDropoutV2: Optimizing Information Flow in Multi-Agent Systems via Test-Time Rectify-or-Reject Pruning

Model ReleasesDGX agent

arXiv:2602.23258v2 Announce Type: replace Abstract: While Multi-Agent Systems (MAS) excel in complex reasoning, they suffer from the cascading impact of erroneous information from individual agents. C

AgentSchool: An LLM-Powered Multi-Agent Simulation for Education

AgentsDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.30144v1 Announce Type: new Abstract: Despite the rapid deployment of LLMs into classrooms, validating educational AI remains uniquely intractable: interventions act on developing learners w

Aggregate Models, Not Explanations: Improving Feature Importance Estimation

ResearchDGX agent

arXiv:2602.11760v2 Announce Type: replace-cross Abstract: Feature-importance methods show promise in transforming machine learning models from predictive engines into tools for scientific discovery. H

Agora: Toward Autonomous Bug Detection in Production-Level Consensus Protocols with LLM Agents

SafetyDGX agent

arXiv:2605.29910v1 Announce Type: cross Abstract: Consensus protocols form the backbone of distributed systems and blockchains, where implementation bugs can cause data corruption and financial losses

AIRGuard: Guarding Agent Actions with Runtime Authority Control

SafetyDGX agent

arXiv:2605.28914v1 Announce Type: cross Abstract: Tool-using language agents turn model decisions into external side effects: they read files, run scripts, call APIs, send messages, and invoke Model C

Aligned but Fragile: Enhancing LLM Safety Robustness via Zeroth-Order Optimization

Model ReleasesDGX agent

arXiv:2605.29396v1 Announce Type: new Abstract: Safety alignment for large language models (LLMs) aims to reduce harmful or unsafe behavior while preserving general utility. However, recent findings r

Alignment-Guided Score Matching for Text-to-Image Alignment in Diffusion Models

Model ReleasesDGX agent

arXiv:2605.30038v1 Announce Type: cross Abstract: Diffusion models generate highly realistic images but often struggle with precise text-image alignment. While recent post-training methods improve ali

AlignVid: Training-Free Attention Scaling for Semantic Fidelity in Text-Guided Image-to-Video Generation

Model ReleasesDGX agent

arXiv:2512.01334v2 Announce Type: replace Abstract: Text-guided image-to-video generation has made substantial progress, yet it still struggles to execute text-specified edits that require substantial

AliMark: Enhancing Robustness of Sentence-Level Watermarking Against Text Paraphrasing

SafetyDGX agent

arXiv:2605.29434v1 Announce Type: cross Abstract: Existing sentence-level watermarking methods enhance robustness to paraphrasing by anchoring watermarks in sentence semantics. However, their prefix-b

Ambient-robust Inverse Rendering using Active RGB-NIR Imaging

ResearchDGX agent

arXiv:2605.30250v1 Announce Type: new Abstract: Inverse rendering aims to reconstruct geometry and reflectance of objects from images. Despite recent progress, existing methods often produces inaccura

AMDP: Asynchronous Multi-Directional Pipeline Parallelism for Large-Scale Models Training

Model ReleasesDGX agent

arXiv:2605.29664v1 Announce Type: cross Abstract: Pipeline parallelism is essential for large-scale model training, but existing asynchronous approaches often degrade convergence due to parameter mism

An accuracy-aware extension to LRP-based pruning for CNNs to prevent cascading accuracy degradation in data-scarce transfer learning

ResearchDGX agent

arXiv:2511.10861v3 Announce Type: replace-cross Abstract: Convolutional Neural Networks (CNNs) pre-trained on large-scale datasets such as ImageNet are widely used as feature extractors to construct h

An Approach for Thyroid Nodule Analysis Using Thermographic Images

AgentsDGX agent

arXiv:2605.29221v1 Announce Type: new Abstract: Thyroid cancer is said to be the second most common type of cancer in female individuals and the third in males by 2030, according to projections. In ge

An End-to-End PyTorch Interface for Differentiable PDE Solvers: A RANS Model-Correction Study

Model ReleasesDGX agent

arXiv:2605.28858v1 Announce Type: cross Abstract: This work presents an end-to-end strategy for solving inverse problems constrained by Partial Differential Equations within a fully differentiable Mac

Analyzing Persona Effects in Generated Explanations from Multimodal LLM Agents in Urban Perception

ResearchDGX agent

arXiv:2605.29064v1 Announce Type: new Abstract: We study how persona prompting shapes language generated by multimodal large language models in an urban perception setting. Using 59,808 annotations fr

Anchorless Diversification for Parallel LLM Ideation

ResearchDGX agent

arXiv:2605.30150v1 Announce Type: new Abstract: LLMs are increasingly used to generate candidate-idea pools for creative tasks where broad exploration is valuable. Parallel inference can be attractive

AnomalyAgent: Training-Free Agentic Models for Zero-/Few-Shot Anomaly Detection

AgentsDGX agent

arXiv:2605.30140v1 Announce Type: new Abstract: Benefiting from generalizability of vision-language models (VLMs) such as CLIP, many zero-/few-shot anomaly detection (AD) approaches have achieved impr

Anti Mode-Collapse in Mean-Field Transformer via Auxiliary Variables

ResearchDGX agent

arXiv:2605.30229v1 Announce Type: new Abstract: We use a mean-field-based transformer model to theoretically investigate how auxiliary variables, such as positional encoding, prevent mode collapse of

AnyMo: Scaling Any-Modality Conditional Motion Generation with Masked Modeling

ResearchDGX agent

arXiv:2605.29488v1 Announce Type: cross Abstract: Conditional human motion generation remains a fundamental challenge in computer vision and robotics. Despite significant progress, current methods are

Anytime-Valid Federated Conformal RAG for LLM Swarms

SafetyDGX agent

arXiv:2605.29139v1 Announce Type: cross Abstract: Federated Conformal RAG (FC-RAG) provides distribution-free coverage for a bandwidth-limited swarm of weak language models, but only at a fixed horizo

Apertus LLM Family Expansion via Distillation and Quantization

Model ReleasesDGX agent

arXiv:2605.29128v1 Announce Type: new Abstract: The wide adoption of LLMs has led to their use in great variety of applications and scenarios, such as chatbot assistants and data annotation, creating

Approximate Proportionality in Online Fair Division

ResearchDGX agent

arXiv:2508.03253v2 Announce Type: replace-cross Abstract: We study the online fair division problem, where indivisible goods arrive sequentially and must be allocated immediately and irrevocably. Prio

Architecture-Sensitive Supervised Fine-Tuning for Screen-Conditioned Action Prediction: A PiSAR Benchmark

Model ReleasesDGX agent

arXiv:2605.29400v1 Announce Type: new Abstract: We benchmark three supervised fine-tuned models against frontier zero-shot baselines on a 661-row held-out slice of PiSAR (Persona, intent, Screen, Acti

Archon: A Unified Multimodal Model for Holistic Digital Human Generation

ResearchDGX agent

arXiv:2605.30311v1 Announce Type: cross Abstract: Digital humans are fundamental to immersive interaction, yet creating a unified model for holistic modalities, including text, audio, motion, and visu

Are LLMs Socially Adaptive? Contrasting Belief Evolution in Large Language Models and Humans

Model ReleasesDGX agent

arXiv:2410.10398v3 Announce Type: replace-cross Abstract: As large language models (LLMs) increasingly engage in complex social interactions, ensuring that their behaviors align with human ethical pri

Aryabhata 2: Scaling Reinforcement Learning for Advanced STEM Reasoning

ApplicationsDGX agent

arXiv:2605.28829v1 Announce Type: cross Abstract: Competitive STEM examinations such as JEE and NEET require multi-step symbolic reasoning, precise numerical computation, and deep conceptual understan

Assessing Dutch Syllabification Algorithms and Improving Accuracy by Combining Phonetic and Orthographic Information through Deep Learning

ResearchDGX agent

arXiv:2605.28834v1 Announce Type: cross Abstract: Syllabification describes the task of dividing words into syllables. Due to many rules and exceptions, training an algorithm to perform syllabificatio

AsymVLM: Asymmetric Token Pruning for Efficient Vision-Language Model Inference

Local AiDGX agent

arXiv:2605.29535v1 Announce Type: new Abstract: Vision-Language Models (VLMs) process thousands of visual tokens per image alongside comparatively few text tokens, yet existing compression methods tre

AtomWorld: A Benchmark for Evaluating Spatial Reasoning in Large Language Models on Crystalline Materials

Model ReleasesDGX agent

arXiv:2510.04704v4 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown promising potential in scientific research, enabling tasks ranging from knowledge retrieval to propert

Attention as In-Context Empirical Bayes: A Two-Stage View via Particle Dynamics

ResearchDGX agent

arXiv:2605.29351v1 Announce Type: new Abstract: We study minimal attention-only transformers under all-token corruption and show they admit a two-stage empirical Bayes interpretation. A single attenti

Attention Asymmetry in AI Layoff Discourse on X: A Computational Analysis of Capital vs Labour Amplification

ResearchDGX agent

arXiv:2605.29367v1 Announce Type: new Abstract: When workers lose jobs to AI-driven restructuring, two very different conversations happen on X (formerly Twitter) at the same time. Tech executives and

AttuneBench: A Conversation-Based Benchmark for LLM Emotional Intelligence

Model ReleasesDGX agent

arXiv:2605.21739v2 Announce Type: replace Abstract: Emotional intelligence (EI), the ability to perceive, understand, and respond appropriately to others' emotional states, is central to human communi

Audio Deepfake Detection with Half-Truth Localisation Using Cross-Attentive Feature Fusion

Model ReleasesDGX agent

arXiv:2605.29531v1 Announce Type: cross Abstract: Audio deepfake detection is well-studied as a binary problem, but partially manipulated speech, where a short synthesised segment is spliced into an o

Audio Jailbreaks in Large Audio-Language Models: Taxonomy, Attack-Defense Analysis, and Cost-Aware Evaluation

SafetyDGX agent

arXiv:2605.30031v1 Announce Type: cross Abstract: Large Audio Language Models (LALMs) expand jailbreak risks from token-level prompting to the full speech perception-to-reasoning pipeline, where unsaf

Auditing Training Data in Generative Music Models via Black-Box Membership Inference

SafetyDGX agent

arXiv:2605.29202v1 Announce Type: new Abstract: Recent advances in text-to-music generation enable high-fidelity synthesis of structured musical audio, raising growing concerns about data provenance,

Auditing Training-Free 3D Shape Retrieval with Diffused Geodesic Moments

Model ReleasesDGX agent

arXiv:2605.29004v1 Announce Type: new Abstract: Reported retrieval scores for training-free shape descriptors conflate local signal design, normalization, aggregation, codebook fitting, and metric cho

Automating Low-Risk Code Review at Meta: RADAR, Risk Calibration, and Review Efficiency

SafetyDGX agent

arXiv:2605.30208v1 Announce Type: cross Abstract: AI-assisted coding tools have altered software production. At Meta, significant lines of code per human-landed diff grew by 105.9% year over year and

AutoSizer: Automatic Sizing of Analog and Mixed-Signal Circuits via Large Language Model (LLM) Agents

Model ReleasesDGX agent

arXiv:2602.02849v2 Announce Type: replace Abstract: The design of Analog and Mixed-Signal (AMS) integrated circuits remains heavily reliant on expert knowledge, with transistor sizing a major bottlene

BadBlocks: Low-Cost and Stealthy Backdoor Attacks Tailored for Text-to-Image Diffusion Models

HardwareDGX agent

arXiv:2508.03221v5 Announce Type: replace-cross Abstract: Despite the remarkable progress of diffusion models in image generation, recent studies reveal their vulnerability to backdoor attacks via cov

Balancing Multimodal Learning through Label Space Reshaping

Model ReleasesDGX agent

arXiv:2605.28869v1 Announce Type: cross Abstract: Multimodal learning often suffers from modality imbalance, where modalities that converge faster dominate optimization while others remain undertraine

Bandit Algorithms for Deep Brain Stimulation

Model ReleasesDGX agent

arXiv:2601.12699v2 Announce Type: replace Abstract: Deep Brain Stimulation (DBS) is an effective treatment for Parkinson's disease, but conventional fixed-parameter stimulation can reduce battery life

Bastion: Budget-Aware Speculative Decoding with Tree-structured Block Diffusion Drafting

HardwareDGX agent

arXiv:2605.29727v1 Announce Type: new Abstract: Block-diffusion drafters have recently emerged as a powerful alternative for speculative decoding by predicting multiple future-token distributions in a

Battery-Sim-Agent: Leveraging LLM-Agent for Inverse Battery Parameter Estimation

Model ReleasesDGX agent

arXiv:2605.29560v1 Announce Type: new Abstract: Parameterizing high-fidelity 'digital twins' of batteries is a critical yet challenging inverse problem that hinders the pace of battery innovation. Pre

Bayesian model selection and misspecification testing in imaging inverse problems only from noisy and partial measurements

ResearchDGX agent

arXiv:2510.27663v3 Announce Type: replace-cross Abstract: Modern imaging techniques heavily rely on Bayesian statistical models to address difficult image reconstruction and restoration tasks. This pa

'Be My Cheese?': Cultural Nuance Benchmarking for Machine Translation in Multilingual LLMs

Model ReleasesDGX agent

arXiv:2602.04729v2 Announce Type: replace Abstract: We present a large-scale human evaluation benchmark for assessing cultural localisation in machine translation produced by state-of-the-art multilin

BEAMS: Benchmarking and Evaluating AI for Modeling and Simulation

SafetyDGX agent

arXiv:2605.28994v1 Announce Type: new Abstract: AI tools to support real world decision making must be able to build simulation models that inform their recommendations and render them interpretable.

Before the Shutter: Aesthetic and Actionable Portrait Photography Planning in 3D Scenes

ApplicationsDGX agent

arXiv:2605.30318v1 Announce Type: cross Abstract: Portrait photography is largely decided before the shutter opens: the subject's pose, the camera configuration, and the lighting devices must be coord

Behavior-Aware Auxiliary Corrections for Off-Policy Temporal-Difference Prediction

Local AiDGX agent

arXiv:2605.28855v1 Announce Type: new Abstract: Temporal-difference learning with function approximation can be unstable under off-policy sampling. TDC stabilizes off-policy TD through an auxiliary co

Behavior-Induced Mirror-Prox Temporal-Difference Learning for Faster Off-Policy Prediction

SafetyDGX agent

arXiv:2605.28849v1 Announce Type: new Abstract: Gradient temporal-difference methods provide stable off-policy prediction with linear function approximation, but their practical performance is strongl

Benchmarking at the Edge of Comprehension

Local AiDGX agent

arXiv:2602.14307v3 Announce Type: replace Abstract: As frontier Large Language Models (LLMs) increasingly saturate new benchmarks shortly after they are published, benchmarking itself is at a juncture

Benchmarking Large Vision-Language Models on CFMME: A Comprehensive Chinese Financial Multimodal Evaluation Dataset

Model ReleasesDGX agent

arXiv:2605.29462v1 Announce Type: cross Abstract: The emergence of Large Vision-Language Models (LVLMs) has substantially expanded model capabilities beyond text-only understanding, enabling unified i

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting

Model ReleasesDGX agent

arXiv:2509.23571v3 Announce Type: replace-cross Abstract: As cyber threats continue to grow in scale and sophistication, blue team defenders increasingly require advanced tools to proactively detect a

Benchmarking Open-Source Safety Guard Models: A Comprehensive Evaluation

Model ReleasesDGX agent

arXiv:2605.28830v1 Announce Type: cross Abstract: As Large Language Models (LLMs) are increasingly deployed in safety-critical applications, robust content moderation becomes essential. We present a c

Benchmarking Positional Encoding Strategies for Transformer-Based EEG Foundation Models

Model ReleasesDGX agent

arXiv:2605.29754v1 Announce Type: new Abstract: Electroencephalography (EEG) is a widely used non-invasive technique for measuring brain activity in brain-computer interface (BCI) applications. Superv

Benchmarking Single-Factor Physical Video-to-Audio Generation

Model ReleasesDGX agent

arXiv:2605.30339v1 Announce Type: new Abstract: Generative video-to-audio (V2A) models produce highly plausible soundtracks, but it remains unclear whether they capture the underlying physical process

BenchTrace: A Benchmark for Testing Reflection Ability and Controlled Evolution in LLM Agents

Model ReleasesDGX agent

arXiv:2605.29225v1 Announce Type: new Abstract: Self-evolving agents improve over time by reflecting on past failures, but existing evaluation is limited in two ways: it measures only task scores, lea

Better Later Than Sooner: Neuro-Symbolic Knowledge Graph Construction via Ontology-grounded Post-extraction Correction

ResearchDGX agent

arXiv:2605.29168v1 Announce Type: new Abstract: Question answering (QA) is a core challenge in AI, particularly for complex queries requiring multi-hop reasoning across documents, or symbolic operatio

Beyond 3D VQAs: Injecting 3D Spatial Priors into Vision-Language Models for Enhanced Geometric Reasoning

ResearchDGX agent

arXiv:2605.30231v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) often struggle with robust 3D spatial reasoning. Prevailing methods that rely on fine-tuning with 3D visual question-ans

Beyond Accuracy: Are Time Series Foundation Models Well-Calibrated?

ResearchDGX agent

arXiv:2510.16060v2 Announce Type: replace-cross Abstract: The recent development of foundation models for time series data has generated considerable interest in using such models across a variety of

← Previous
1…550551552553554…1049
Next →