AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
Human
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
21 Apr 2026

AVRT: Audio-Visual Reasoning Transfer through Single-Modality Teachers

Model ReleasesDGX agent

arXiv:2604.16617v1 Announce Type: new Abstract: Recent advances in reasoning models have shown remarkable progress in text-based domains, but transferring those capabilities to multimodal settings, e.

AWPD: Frequency Shield Network for Agnostic Watermark Presence Detection

Model ReleasesDGX agent

arXiv:2603.06723v3 Announce Type: replace Abstract: Invisible watermarks, as an essential technology for image copyright protection, have been widely deployed with the rapid development of social medi

Back into Plato's Cave: Examining Cross-modal Representational Convergence at Scale

SafetyDGX agent

arXiv:2604.18572v1 Announce Type: new Abstract: The Platonic Representation Hypothesis suggests that neural networks trained on different modalities (e.g., text and images) align and eventually conver

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Back to Repair: A Minimal Denoising Network for Time Series Anomaly Detection

Model ReleasesDGX agent

arXiv:2604.17388v1 Announce Type: new Abstract: We introduce JuRe (Just Repair), a minimal denoising network for time series anomaly detection that exposes a central finding: architectural complexity

Balance-Guided Sparse Identification of Multiscale Nonlinear PDEs with Small-coefficient Terms

ResearchDGX agent

arXiv:2604.18414v1 Announce Type: new Abstract: Data-driven discovery of governing equations has advanced significantly in recent years; however, existing methods often struggle in multiscale systems

Balanced Co-Clustering of Users and Items for Embedding Table Compression in Recommender Systems

Model ReleasesDGX agent

arXiv:2604.18351v1 Announce Type: cross Abstract: Recommender systems have advanced markedly over the past decade by transforming each user/item into a dense embedding vector with deep learning models

BARD: Bridging AutoRegressive and Diffusion Vision-Language Models Via Highly Efficient Progressive Block Merging and Stage-Wise Distillation

ResearchDGX agent

arXiv:2604.16514v1 Announce Type: new Abstract: Autoregressive vision-language models (VLMs) deliver strong multimodal capability, but their token-by-token decoding imposes a fundamental inference bot

Barrier-enforced multi-objective optimization for direct point and sharp interval forecasting

ResearchDGX agent

arXiv:2604.18492v1 Announce Type: new Abstract: This paper proposes a multi-step probabilistic forecasting framework using a single neural-network based model to generate simultaneous point and interv

BASIL: Bayesian Assessment of Sycophancy in LLMs

ApplicationsDGX agent

arXiv:2508.16846v5 Announce Type: replace-cross Abstract: Sycophancy (overly agreeable or flattering behavior) poses a fundamental challenge for human-AI collaboration, particularly in high-stakes dec

BASIS: Balanced Activation Sketching with Invariant Scalars for 'Ghost Backpropagation'

SafetyDGX agent

arXiv:2604.16324v1 Announce Type: new Abstract: The activation memory required for exact backpropagation scales linearly with network depth, context length, and feature dimensionality, forming an O(L

BasketHAR: A Multimodal Dataset for Human Activity Recognition and Sport Analysis in Basketball Training Scenarios

Model ReleasesDGX agent

arXiv:2604.17065v1 Announce Type: new Abstract: Human Activity Recognition (HAR) involves the automatic identification of user activities and has gained significant research interest due to its broad

Batch-Adaptive Causal Annotations

SafetyDGX agent

arXiv:2502.10605v3 Announce Type: replace-cross Abstract: Estimating the causal effects of interventions is crucial to policy and decision-making, yet outcome data are often missing or subject to non-

Bayesian Neural Networks: An Introduction and Survey

ResearchDGX agent

arXiv:2006.12024v3 Announce Type: replace-cross Abstract: Neural Networks (NNs) have provided state-of-the-art results for many challenging machine learning tasks such as detection, regression and cla

Beam-Plasma Collective Oscillations in Intense Charged-Particle Beams: Dielectric Response Theory, Langmuir Wave Dispersion, and Unsupervised Detection via Prometheus

ResearchDGX agent

arXiv:2603.10457v3 Announce Type: replace-cross Abstract: We develop a theoretical and computational framework for beam-plasma collective oscillations in intense charged-particle beams at intermediate

Before You Interpret the Profile: Validity Scaling for LLM Metacognitive Self-Report

ResearchDGX agent

arXiv:2604.17707v1 Announce Type: new Abstract: Clinical personality assessment screens response validity before interpreting substantive scales. LLM evaluation does not. We apply the validity scaling

BEFT: Bias-Efficient Fine-Tuning of Language Models in Low-Data Regimes

Model ReleasesDGX agent

arXiv:2509.15974v2 Announce Type: replace Abstract: Fine-tuning the bias terms of large language models (LLMs) has the potential to achieve unprecedented parameter efficiency while maintaining competi

BenchMarker: An Education-Inspired Toolkit for Highlighting Flaws in Multiple-Choice Benchmarks

Model ReleasesDGX agent

arXiv:2602.06221v2 Announce Type: replace Abstract: Multiple-choice question answering (MCQA) is standard in NLP, but benchmarks lack rigorous quality control. We present BenchMarker, an education-ins

Benchmarking Real-Time Question Answering via Executable Code Workflows

Model ReleasesDGX agent

arXiv:2604.16349v1 Announce Type: cross Abstract: Retrieving real-time information is a fundamental capability for search-integrated agents in real-world applications. However, existing benchmarks are

Benchmarking System Dynamics AI Assistants: Cloud Versus Local LLMs on CLD Extraction and Discussion

Model ReleasesDGX agent

arXiv:2604.18566v1 Announce Type: cross Abstract: We present a systematic evaluation of large language model families -- spanning both proprietary cloud APIs and locally-hosted open-source models -- o

BengaliMoralBench: A Benchmark for Auditing Moral Reasoning in Large Language Models within Bengali Language and Culture

Model ReleasesDGX agent

arXiv:2511.03180v2 Announce Type: replace Abstract: As multilingual Large Language Models (LLMs) gain traction across South Asia, their alignment with local ethical norms, particularly for Bengali, sp

Better with Less: Tackling Heterogeneous Multi-Modal Image Joint Pretraining via Conditioned and Degraded Masked Autoencoder

SafetyDGX agent

arXiv:2604.16952v1 Announce Type: new Abstract: Learning robust representations across extremely heterogeneous modalities remains a fundamental challenge in multi-modal vision. As a critical and profo

Beyond Attack Success Rate: A Multi-Metric Evaluation of Adversarial Transferability in Medical Imaging Models

TutorialsDGX agent

arXiv:2604.16532v1 Announce Type: new Abstract: While deep learning systems are becoming increasingly prevalent in medical image analysis, their vulnerabilities to adversarial perturbations raise seri

Beyond Binary Contrast: Modeling Continuous Skeleton Action Spaces with Transitional Anchors

ResearchDGX agent

arXiv:2604.17914v1 Announce Type: new Abstract: Self-supervised contrastive learning has emerged as a powerful paradigm for skeleton-based action recognition by enforcing consistency in the embedding

Beyond Black-Box Labels: Interpretable Criteria for Diagnosing SubjectiveNLP Tasks

ResearchDGX agent

arXiv:2604.17022v1 Announce Type: new Abstract: Subjective NLP datasets typically aggregate annotator judgments into a single gold label, making it difficult to diagnose whether disagreement reflects

Beyond Feature Fusion: Contextual Bayesian PEFT for Multimodal Uncertainty Estimation

Model ReleasesDGX agent

arXiv:2604.16615v1 Announce Type: new Abstract: We introduce CoCo-LoRA, a multimodal, uncertainty-aware parameter-efficient fine-tuning method for text prediction tasks accompanied by audio context. E

Beyond Fine-Tuning: In-Context Learning and Chain-of-Thought for Reasoned Distractor Generation

ResearchDGX agent

arXiv:2604.17574v1 Announce Type: new Abstract: Distractor generation (DG) remains a labor-intensive task that still significantly depends on domain experts. The task focuses on generating plausible y

Beyond 'I Don't Know': Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty

Model ReleasesDGX agent

arXiv:2604.17293v1 Announce Type: new Abstract: Reliable Large Language Models (LLMs) should abstain when confidence is insufficient. However, prior studies often treat refusal as a generic 'I don't k

Beyond Memorization: Extending Reasoning Depth with Recurrence, Memory and Test-Time Compute Scaling

Local AiDGX agent

arXiv:2508.16745v2 Announce Type: replace Abstract: Reasoning is a core capability of large language models, yet how multi-step reasoning is learned and executed remains unclear. We study this questio

Beyond Overlap Metrics: Rewarding Reasoning and Preferences for Faithful Multi-Role Dialogue Summarization

SafetyDGX agent

arXiv:2604.17188v1 Announce Type: new Abstract: Multi-role dialogue summarization requires modeling complex interactions among multiple speakers while preserving role-specific information and factual

Beyond Pattern Matching: Seven Cross-Domain Techniques for Prompt Injection Detection

Model ReleasesDGX agent

arXiv:2604.18248v1 Announce Type: cross Abstract: Current open-source prompt-injection detectors converge on two architectural choices: regular-expression pattern matching and fine-tuned transformer c

Beyond Reproduction: A Paired-Task Framework for Assessing LLM Comprehension and Creativity in Literary Translation

Model ReleasesDGX agent

arXiv:2604.18169v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for creative tasks such as literary translation. Yet translational creativity remains underexplored a

Beyond Static Benchmarks: Synthesizing Harmful Content via Persona-based Simulation for Robust Evaluation

ResearchDGX agent

arXiv:2604.17020v1 Announce Type: new Abstract: Static benchmarks for harmful content detection face limitations in scalability and diversity, and may also be affected by contamination from web-scale

Beyond the Failures: Rethinking Foundation Models in Pathology

ResearchDGX agent

arXiv:2510.23807v5 Announce Type: replace-cross Abstract: Despite their successes in vision and language, foundation models have stumbled in pathology, revealing low accuracy, instability, and heavy c

Beyond URLs: Metadata Diversity and Position for Efficient LLM Pretraining

ResearchDGX agent

arXiv:2511.21613v2 Announce Type: replace Abstract: Incorporating metadata in Large Language Models (LLMs) pretraining has recently emerged as a promising approach to accelerate training. However prio

Beyond Verifiable Rewards: Rubric-Based GRM for Reinforced Fine-Tuning SWE Agents

ResearchDGX agent

arXiv:2604.16335v1 Announce Type: new Abstract: Despite recent progress in Large Language Model (LLM) Agents for Software Engineering (SWE) tasks, end-to-end fine-tuning typically relies on verifiable

Beyond Word Boundaries: A Hebrew Coreference Benchmark and an Evaluation Protocol for Morphologically Complex Text

Model ReleasesDGX agent

arXiv:2604.17108v1 Announce Type: new Abstract: Coreference Resolution (CR) is a fundamental NLP task critical for long-form tasks as information extraction, summarization, and many business applicati

BhashaSutra: A Task-Centric Unified Survey of Indian NLP Datasets, Corpora, and Resources

ResearchDGX agent

arXiv:2604.18423v1 Announce Type: new Abstract: India's linguistic landscape, spanning 22 scheduled languages and hundreds of marginalized dialects, has driven rapid growth in NLP datasets, benchmarks

Bi-LoRA: Efficient Sharpness-Aware Minimization for Fine-Tuning Large-Scale Models

Model ReleasesDGX agent

arXiv:2508.19564v2 Announce Type: replace Abstract: Fine-tuning large-scale pre-trained models with limited data presents significant challenges for generalization. While Sharpness-Aware Minimization

Bias-constrained multimodal intelligence for equitable and reliable clinical AI

SafetyDGX agent

arXiv:2604.16884v1 Announce Type: new Abstract: The integration of medical imaging and clinical text has enabled the emergence of generalist artificial intelligence (AI) systems for healthcare. Howeve

BIASEDTALES-ML: A Multilingual Dataset for Analyzing Narrative Attribute Distributions in LLM-Generated Stories

SafetyDGX agent

arXiv:2604.17008v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used to generate narrative content, including children's stories, which play an important role in social a

Bielik Guard: Efficient Polish Language Safety Classifiers for LLM Content Moderation

Model ReleasesDGX agent

arXiv:2602.07954v4 Announce Type: replace Abstract: As Large Language Models (LLMs) become increasingly deployed in Polish language applications, the need for efficient and accurate content safety cla

Bilinear Input Modulation for Mamba: Koopman Bilinear Forms for Memory Retention and Multiplicative Computation

ResearchDGX agent

arXiv:2604.17221v1 Announce Type: cross Abstract: Selective State Space Models (SSMs), notably Mamba, employ diagonal state transitions that limit both memory retention and bilinear computational capa

BioVLM: Routing Prompts, Not Parameters, for Cross-Modality Generalization in Biomedical VLMs

ResearchDGX agent

arXiv:2604.17629v1 Announce Type: new Abstract: Pretrained biomedical vision-language models (VLMs) such as BioMedCLIP perform well on average but often degrade on challenging modalities where inter-c

Bit-Flip Vulnerability of Shared KV-Cache Blocks in LLM Serving Systems

HardwareDGX agent

arXiv:2604.17249v1 Announce Type: cross Abstract: Rowhammer on GPU DRAM has enabled adversarial bit flips in model weights; shared KV-cache blocks in LLM serving systems present an analogous but previ

Block-encodings as programming abstractions: The Eclipse Qrisp BlockEncoding Interface

TutorialsDGX agent

arXiv:2604.18276v1 Announce Type: cross Abstract: Block-encoding is a foundational technique in modern quantum algorithms, enabling the implementation of non-unitary operations by embedding them into

BOIL: Learning Environment Personalized Information

AgentsDGX agent

arXiv:2604.17137v1 Announce Type: new Abstract: Navigating complex environments poses challenges for multi-agent systems, requiring efficient extraction of insights from limited information. In this p

Boltzmann Machine Learning with a Parallel, Persistent Markov chain Monte Carlo method for Estimating Evolutionary Fields and Couplings from a Protein Multiple Sequence Alignment

SafetyDGX agent

arXiv:2604.18022v1 Announce Type: cross Abstract: The inverse Potts problem for estimating evolutionary single-site fields and pairwise couplings in homologous protein sequences from their single-site

Bolzano: Case Studies in LLM-Assisted Mathematical Research

AgentsDGX agent

arXiv:2604.16989v1 Announce Type: new Abstract: We report new results on six problems in mathematics and theoretical computer science, produced with the assistance of Bolzano, an open-source multi-age

BOOKAGENT: Orchestrating Safety-Aware Visual Narratives via Multi-Agent Cognitive Calibration

Model ReleasesDGX agent

arXiv:2604.16541v1 Announce Type: new Abstract: Recent advancements in Large Generative Models (LGMs) have revolutionized multi-modal generation. However, generating illustrated storybooks remains an

BOP-ASK: Object-Interaction Reasoning for Vision-Language Models

Model ReleasesDGX agent

arXiv:2511.16857v3 Announce Type: replace Abstract: Vision Language Models (VLMs) have achieved impressive performance on spatial reasoning benchmarks, yet these evaluations mask critical weaknesses i

Bounded Graph Clustering with Graph Neural Networks

ResearchDGX agent

arXiv:2512.05623v2 Announce Type: replace Abstract: In community detection, many methods require the user to specify the number of clusters in advance since an exhaustive search over all possible valu

Bounded Ratio Reinforcement Learning

SafetyDGX agent

arXiv:2604.18578v1 Announce Type: new Abstract: Proximal Policy Optimization (PPO) has become the predominant algorithm for on-policy reinforcement learning due to its scalability and empirical robust

Brain-CLIPLM: Decoding Compressed Semantic Representations in EEG for Language Reconstruction

ResearchDGX agent

arXiv:2604.16370v1 Announce Type: new Abstract: Decoding natural language from non-invasive electroencephalography (EEG) remains fundamentally limited by low signal-to-noise ratio and restricted infor

Brain-Inspired Capture: Evidence-Driven Neuromimetic Perceptual Simulation for Visual Decoding

Model ReleasesDGX agent

arXiv:2604.17927v1 Announce Type: new Abstract: Visual decoding of neurophysiological signals is a critical challenge for brain-computer interfaces (BCIs) and computational neuroscience. However, curr

BrainMem: Brain-Inspired Evolving Memory for Embodied Agent Task Planning

AgentsDGX agent

arXiv:2604.16331v1 Announce Type: cross Abstract: Embodied task planning requires agents to execute long-horizon, goal-directed actions in complex 3D environments, where success depends on both immedi

Bridge-Centered Metapath Classification Using R-GCN-VGAE for Disaster-Resilient Maintenance Decisions

ApplicationsDGX agent

arXiv:2604.18399v1 Announce Type: new Abstract: Daily infrastructure management in preparation for disasters is critical for urban resilience. When bridges remain resilient against disaster-induced ex

BRIDGE the Gap: Mitigating Bias Amplification in Automated Scoring of English Language Learners via Inter-group Data Augmentation

SafetyDGX agent

arXiv:2602.23580v2 Announce Type: replace Abstract: In the field of educational assessment, automated scoring systems increasingly rely on deep learning and large language models (LLMs). However, thes

BridgeEQA: Virtual Embodied Agents for Real Bridge Inspections

Model ReleasesDGX agent

arXiv:2511.12676v2 Announce Type: replace Abstract: Deploying embodied agents that can answer questions about their surroundings in realistic real-world settings remains difficult, partly due to the s

Bridging Coarse and Fine Recognition: A Hybrid Approach for Open-Ended Multi-Granularity Object Recognition in Interactive Educational Games

ResearchDGX agent

arXiv:2604.16785v1 Announce Type: new Abstract: Recent advances in Multimodal Large Language Models (MLLMs) have enabled open-ended object recognition, yet they struggle with fine-grained tasks. In co

Bridging the Culture Gap: A Framework for LLM-Driven Socio-Cultural Localization of Math Word Problems in Low-Resource Languages

Local AiDGX agent

arXiv:2508.14913v4 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated significant capabilities in solving mathematical problems expressed in natural language. However, mul

← Previous
1…882883884885886…998
Next →