AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,259
  • Agents7,702
  • Applications5,506
  • Concepts5
  • Hardware1,896
  • Industry6,187
  • Local Ai5,048
  • Model Releases24,516
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,677
  • Tutorials3,454

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,259
  • Agents7,702
  • Applications5,506
  • Concepts5
  • Hardware1,896
  • Industry6,187
  • Local Ai5,048
  • Model Releases24,516
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,677
  • Tutorials3,454

Source
Human
90,259Total entries
1Added by human
90,258Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
90,258 results
28 May 2026

Bayesian Optimization Parameter Tuning Framework for a Lyapunov Based Path Following Controller

Model ReleasesDGX agent

arXiv:2512.12649v2 Announce Type: replace Abstract: Parameter tuning in real-world experiments is constrained by the limited evaluation budget available on hardware. The path-following controller stud

BEAR: Budgeted Evidence Allocation for Multi-Document Reasoning

ApplicationsDGX agent

arXiv:2601.18116v2 Announce Type: replace Abstract: We argue that multi-document reasoning is constrained not only by how much text a model can read, but also by how limited query-time evidence budget

Becoming an AI-native company is existential for every business, says Dell CMO

SafetyDGX agent

Becoming an AI-native company is no longer the competitive advantage it was five minutes ago. Now, it’s an empirical obligation, and Dell Technologies Inc. made that case to customers at its annual fl

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Behavioural Analysis of Alignment Faking

SafetyDGX agent

arXiv:2605.27681v1 Announce Type: new Abstract: Alignment faking (AF) refers to a model strategically complying with a training objective to avoid behavioural modification while preserving its deploym

Benchmarking AI for low-resource contexts: Thinking beyond leaderboards

Model ReleasesDGX agent

arXiv:2605.28508v1 Announce Type: new Abstract: Existing AI evaluation practices often fail to capture how systems actually perform in low-resource environments, where operational constraints shape us

Benchmarking and Mechanistic Analysis of Vision-Language Models for Cross-Depiction Assembly Instruction Alignment

Model ReleasesDGX agent

arXiv:2604.00913v2 Announce Type: replace-cross Abstract: 2D assembly diagrams are often abstract and hard to follow, creating a need for intelligent assistants that can monitor progress, detect error

Benchmarking Fairness in Spiking Neural Networks: Data Bias, Spurious Features, and Hardware Effects

Model ReleasesDGX agent

arXiv:2605.27407v1 Announce Type: cross Abstract: Evaluating fairness in Spiking Neural Networks (SNNs) demands rigorous benchmarks that reflect real-world complexities, yet existing assessments remai

Benchmarking Inductive Biases for Multivariate Time-Series Anomaly Detection with a Robust Multi-View Channel-Graph Detector

Model ReleasesDGX agent

arXiv:2605.28103v1 Announce Type: new Abstract: We present a unified experiment, analysis, and benchmark study of multivariate time-series (MTS) anomaly detection. Ten family-representative detectors

Benchmarking Ultrasound Foundation Models for Fetal Plane Classification

Model ReleasesDGX agent

arXiv:2605.27796v1 Announce Type: cross Abstract: Ultrasound is widely used in obstetric care due to its safety, accessibility, and real-time imaging. However, interpretation remains operator-dependen

Benchmarks are Not Enough: RAMP for Runtime Assessing of Agentic Models in Production Systems

Model ReleasesDGX agent

arXiv:2605.27492v1 Announce Type: cross Abstract: LLM agents are rapidly evolving from coding assistants into autonomous software engineering systems. However, existing evaluation methodologies remain

BenGER: Benchmarking LLM Systems on Subsumption-Based Legal Reasoning in German Law

Model ReleasesDGX agent

arXiv:2605.28183v1 Announce Type: cross Abstract: We introduce the BenGER (Benchmark for German Law) dataset for evaluating LLM systems on subsumption-based legal reasoning in German law. The BenGER d

Better Accuracies, Worse Reasoning: A Step-Level Audit of Medical Chain-of-Thought Distillation

Model ReleasesDGX agent

arXiv:2605.28301v1 Announce Type: new Abstract: Chain-of-thought (CoT) distillation trains a smaller model to imitate a teacher's reasoning trace, but it is typically evaluated by final-answer metrics

Better heads do not guarantee better binarized constituency parsing

ResearchDGX agent

arXiv:2605.28131v1 Announce Type: new Abstract: We revisit punctuation-aware tree binarization for constituency parsing and ask whether dependency-induced headedness improves binary parser supervision

Beyond being fast, LiteParse is designed to provide highly accurate, semantically coherent text for LLM use. We benchmarked every open-sourc…

Model ReleasesDGX agent

Beyond being fast, LiteParse is designed to provide highly accurate, semantically coherent text for LLM use. We benchmarked every open-source, model-free PDF parser on LLM QA tasks - from PyPDF to PyM

Beyond Binary Moral Judgment: Modeling Ethical Pluralism in AI

Model ReleasesDGX agent

arXiv:2605.28707v1 Announce Type: new Abstract: Critical decision-making in socially consequential spaces is increasingly involving AI systems at varying capacities. Yet, despite the ubiquity of auton

Beyond Binary: Sim-to-Real Dexterous Manipulation with Physics-Grounded Contact Representation

SafetyDGX agent

arXiv:2605.28812v1 Announce Type: cross Abstract: A primary bottleneck in contact-rich manipulation is the difficulty of collecting real-world data. Sim-to-real reinforcement learning offers a scalabl

Beyond Chunk-Local Extraction: Cross-Chunk Graph Augmentation for GraphRAG

ResearchDGX agent

arXiv:2605.28004v1 Announce Type: new Abstract: GraphRAG extends retrieval-augmented generation by organizing corpora as explicit knowledge graphs, enabling graph-based retrieval for complex question

Beyond External Monitors: Enhancing Transparency of Large Language Models for Easier Monitoring

Model ReleasesDGX agent

arXiv:2502.05242v3 Announce Type: replace-cross Abstract: Large language models (LLMs) are becoming increasingly capable, but the mechanisms of their thinking and decision-making processes remain uncl

Beyond Input Understanding: Diagnosing Multilingual Mathematical Reasoning with Directed Acyclic Trace Graphs

ResearchDGX agent

arXiv:2605.27715v1 Announce Type: new Abstract: Large reasoning models (LRMs) achieve strong mathematical reasoning performance in English, but remain much less reliable in many low- and medium-resour

Beyond Lipschitz: Data-Driven Robustness via Discrete Modulus of Continuity

Local AiDGX agent

arXiv:2605.28729v1 Announce Type: cross Abstract: Robustness of neural networks is commonly quantified via local or global Lipschitz constants. However, Lipschitz continuity can be overly coarse or ov

Beyond Model Ranking: Predictability-Aligned Evaluation for Time Series Forecasting

Model ReleasesDGX agent

arXiv:2509.23074v3 Announce Type: replace-cross Abstract: In the era of increasingly complex AI models for time series forecasting, progress is often measured by marginal improvements on benchmark lea

Beyond Motion Primitives: Behavioral Activity Recognition from Head-Mounted IMU

Model ReleasesDGX agent

arXiv:2605.27464v1 Announce Type: cross Abstract: AR smart glasses need continuous behavioral context to offer proactive assistance, yet their most practical always-on sensor, the head-mounted Inertia

Beyond One Path: Evaluating and Enhancing Divergent Thinking in Interactive LLM Agents

Model ReleasesDGX agent

arXiv:2605.28465v1 Announce Type: new Abstract: Divergent thinking is a core dimension of creativity, yet existing evaluations of Large Language Models (LLMs) treat them as single-turn text generation

Beyond pass@k: Redundancy-Aware RLVR for Multi-Sample Code Generation

ResearchDGX agent

arXiv:2605.28022v1 Announce Type: new Abstract: LLMs for code generation are commonly evaluated in repeated-sampling settings using Pass@k, where multiple candidate programs are executed against unit

Beyond Surrogate Gradients: Fully Differentiable Token Pruning for Vision-Language Models

ResearchDGX agent

arXiv:2605.28051v1 Announce Type: new Abstract: Visual token pruning reduces the computational cost of Vision-Language Models (VLMs) by removing redundant visual tokens. Existing methods typically rel

Bias Leaves a Gradient Trail: Label-Free Bias Identification via Gradient Probes on Concept Decompositions

Model ReleasesDGX agent

arXiv:2605.28780v1 Announce Type: new Abstract: Vision classifiers can exploit spurious correlations, achieving high in-distribution accuracy yet failing under distribution shift. Existing approaches

BiasEdit: A Training-Free Bias-Detect-and-Edit Framework for Learning Fair Visual Classifiers

SafetyDGX agent

arXiv:2605.28450v1 Announce Type: cross Abstract: Visual data from the Web power image classifiers, which often underpin many web services, such as recommendation and content moderation. However, the

Big migrations and refactors are some of a team's most important work, and the easiest to push off to a 'better time' since they'd tie up en…

Model ReleasesDGX agent

Big migrations and refactors are some of a team's most important work, and the easiest to push off to a 'better time' since they'd tie up engineers for a quarter. With dynamic workflows, Claude can no

Big show tomorrow: 1. Beautiful apps with Canvas, with @stanhuan — explore ideas, design multiple variations, turn a design into a working a…

ToolsDGX agent

Big show tomorrow: 1. Beautiful apps with Canvas, with @stanhuan — explore ideas, design multiple variations, turn a design into a working app, fine-tune with design controls. 2. Customizable sign-in

Bilinear Coordinate Alignment for Training-Free Task-Vector Transfer

Model ReleasesDGX agent

arXiv:2605.28444v1 Announce Type: new Abstract: Fine-tuning large-scale pre-trained models is a recent prevalent paradigm for adapting general representations to specialized tasks. However, when a new

Bio-Inspired Self-Supervised Learning for Wrist-worn Accelerometer Data

ResearchDGX agent

arXiv:2603.10961v2 Announce Type: replace Abstract: Wearable accelerometers enable large-scale health monitoring, yet learning robust human-activity representations has been constrained by scarce labe

BioELX: Cross-lingual Biomedical Entity Linking via Alias-based Retrieval and LLM Ranking

Model ReleasesDGX agent

arXiv:2605.27380v1 Announce Type: cross Abstract: Cross-lingual biomedical entity linking (BEL) maps mentions in any language to unique identifiers in a biomedical knowledge base (KB), supporting clin

BIRDNet: Mining and Encoding Boolean Implication Knowledge Graphs as Interpretable Deep Neural Networks

ResearchDGX agent

arXiv:2605.28739v1 Announce Type: cross Abstract: Tabular data in knowledge-rich domains often carries a latent prior in the form of Boolean implication relationships (BIRs) between pairs of features.

BIRDS: Characterizing and Understanding Biodiversity Impact of Large Language Model Serving

ResearchDGX agent

arXiv:2605.27480v1 Announce Type: cross Abstract: Large language model (LLM) serving creates environmental impacts beyond carbon and water, including ecosystem damage through biodiversity-related path

BlazeEdit: Generalist Image Editing on Mobile Devices with Image-to-Image Diffusion Models

Model ReleasesDGX agent

arXiv:2605.28067v1 Announce Type: new Abstract: The remarkable generation quality of modern diffusion models often comes at the cost of massive parameter counts, which necessitate server-side inferenc

Blind PRNG Hijacking: An Undetectable Integrity-Preserving Attack Against LLM Watermarking

SafetyDGX agent

arXiv:2605.28632v1 Announce Type: cross Abstract: Cryptographic watermarking is a leading defense for attributing text generated by large language models (LLMs). Existing schemes, including KGW, Unigr

Bound-Constrained Sparse Representation for Electrical Impedance Tomography

ResearchDGX agent

arXiv:2605.28392v1 Announce Type: new Abstract: This study proposes a bound-constrained sparse representation (BC-SR) framework for electrical impedance tomography (EIT), aimed at improving conductivi

Boundary Suppression Asymmetry in Post-trained Assistants: Over-expansion as a Controllability Cost

SafetyDGX agent

arXiv:2605.27969v1 Announce Type: new Abstract: Post-trained language-model assistants are often optimized to avoid under-answering, encouraging complete, helpful, cautious, and proactive responses. W

Bounded-Compute Multimodal Regression for Product-Rating Prediction

Model ReleasesDGX agent

arXiv:2605.27737v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly attractive for multimodal quality assessment, but their default reliance on autoregressive text generatio

BPPO: Binary Prefix Policy Optimization for Efficient GRPO-Style Reasoning RL with Concise Responses

SafetyDGX agent

arXiv:2605.28028v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) is widely used for training reasoning models, but updating all sampled completions in each group incurs substa

BREAKING: Anthropic just dropped Opus 4.8—and it is a MONSTER We've been testing for about a week @every and our verdict is they could've ju…

Model ReleasesDGX agent

BREAKING: Anthropic just dropped Opus 4.8—and it is a MONSTER We've been testing for about a week @every and our verdict is they could've just called it Opus 5, it's that good. Here's our vibe check:

Breaking the Script Barrier: Enabling Automatic Alignment for PoS-based ASR Error Analysis in Non-Latin Scripts

SafetyDGX agent

arXiv:2605.28438v1 Announce Type: new Abstract: Automatic Speech Recognition (ASR) systems are commonly evaluated using aggregate metrics such as Word Error Rate (WER), which do not capture the lingui

Bridging Maximum Likelihood and Optimal Transport for Efficient Inference and Model Selection in Stochastic Block Models

ResearchDGX agent

arXiv:2605.28488v1 Announce Type: cross Abstract: We study inference in stochastic block models (SBMs) through the lens of optimal transport (OT). We first establish that maximum likelihood variationa

Bridging the Detection-to-Abstention Gap in Reasoning Models under Insufficient Information

SafetyDGX agent

arXiv:2605.28070v1 Announce Type: new Abstract: We highlight a failure mode of large reasoning models on questions with insufficient information: models may recognize that a problem is under-specified

Bridging the Generalization Gap in Adverse Weather Segmentation: A Training Recipe Perspective

ResearchDGX agent

arXiv:2605.27962v1 Announce Type: new Abstract: This paper describes our approach for the 8th UG2+ Workshop (CVPR 2026) Track~2, which targets semantic segmentation of outdoor scenes degraded by five

Bridging the Sampling Distribution Shift in Radio Map Estimation: A Trajectory-Aware Paradigm

ResearchDGX agent

arXiv:2605.28234v1 Announce Type: new Abstract: Learning-based radio map estimation (RME) plays a critical role in UAV-assisted wireless sensing, enabling tasks such as coverage prediction and network

Bridging the Stability-Expressivity Gap: Synthetic Data Scaling and Preference Alignment for Low-Resource Spoken Language Models

Model ReleasesDGX agent

arXiv:2605.27383v1 Announce Type: cross Abstract: Spoken Language Models (SLMs) have emerged as a promising paradigm for speech synthesis by bypassing explicit grapheme-to-phoneme pipelines. However,

BuddyBench: A Privacy-Constrained Multi-Task Benchmark for Pediatric Social-Communication Personalization

Model ReleasesDGX agent

arXiv:2605.28089v1 Announce Type: new Abstract: BuddyBench introduces a privacy-constrained multi-task benchmark for pediatric social-communication personalization. Unlike existing neurodevelopmental

Build a custom portal with embedded Amazon SageMaker AI MLflow Apps

TutorialsDGX agent

In this post, you learn how to build a custom portal with embedded SageMaker AI MLflow Apps UI. You walk through the architecture pattern behind a React front end paired with a Flask reverse proxy tha

Build a test suite that grows with your agent with dataset management in Amazon Bedrock AgentCore

Model ReleasesDGX agent

Agent evaluation is most powerful when you combine fast-moving online signals with stable offline baselines. To understand whether your agent is truly improving over time, you need a fixed benchmark a

Building a real-time power outage map with Next.js on Vercel

ToolsDGX agent

This article demonstrates how to build a real-time power outage mapping application using Next.js and deploy it on Vercel's platform. It likely covers implementing live data updates, geolocation featu

Building Community-Centred NLP Resources for Puno Quechua

Model ReleasesDGX agent

arXiv:2605.28253v1 Announce Type: new Abstract: The preservation of under-resourced languages requires digital tools and resources shaped by and for their speakers. We present the first dedicated ASR

Bullet Trains: Parallelizing Training of Temporally Precise Spiking Neural Networks

ResearchDGX agent

arXiv:2603.13283v2 Announce Type: replace-cross Abstract: Continuous-time, event-native spiking neural networks (SNNs) operate strictly on spike events, treating spike timing and ordering as the repre

C-MIG: Multi-view Information Gain-based Retrieval-Augmented Generation for Clinical Diagnosis Reasoning

TutorialsDGX agent

arXiv:2605.27860v1 Announce Type: new Abstract: Retrieval-augmented generation combined with reinforcement learning has shown promise for grounding large language models in trustworthy medical evidenc

Calibrated Inference for the Conditional Average Treatment Effect in the Few-Placebo Regime via Gaussian Processes

SafetyDGX agent

arXiv:2605.27473v1 Announce Type: cross Abstract: Estimating how much an intervention helps a given individual the conditional average treatment effect (CATE) is increasingly central to decision-makin

Calibrating Conservatism for Scalable Oversight

AgentsDGX agent

arXiv:2605.28807v1 Announce Type: new Abstract: Agentic AI systems capable of autonomous planning and extended environmental interaction pose a fundamental control problem: how can humans maintain mea

California AG Rob Bonta sues 23andMe, alleging it failed to protect sensitive user data in a 2023 breach that affected ~7M people across the US (Jaimie Ding/Associated Press)

IndustryDGX agent

Jaimie Ding / Associated Press: California AG Rob Bonta sues 23andMe, alleging it failed to protect sensitive user data in a 2023 breach that affected ~7M people across the US — California's attorney

CALM-IT: Generating Realistic Long-Form Motivational Interviewing Dialogues with Dual-Actor Conversational Dynamics Tracking

ResearchDGX agent

arXiv:2601.10085v2 Announce Type: replace Abstract: Therapeutic dialogue is not a sequence of isolated responses: client goals, motivation, resistance, and therapeutic alliance evolve over time. Yet c

CaMBRAIN: Real-time, Continuous EEG Inference with Causal State Space Models

ResearchDGX agent

arXiv:2605.28792v1 Announce Type: new Abstract: Electroencephalography (EEG) is a critical, non-invasive method to monitor electrical brain activity. EEGs can span anywhere from a couple seconds to mu

Camellia: Benchmarking Cultural Biases in LLMs for Asian Languages

Model ReleasesDGX agent

arXiv:2510.05291v2 Announce Type: replace Abstract: As Large Language Models (LLMs) develop stronger multilingual capabilities, their sensitivity to culturally diverse entities becomes increasingly im

← Previous
1…798799800801802…1505
Next →