AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
Research

Bayesian Neural Networks: An Introduction and Survey

DGX agent

arXiv:2006.12024v3 Announce Type: replace-cross Abstract: Neural Networks (NNs) have provided state-of-the-art results for many challenging machine learning tasks such as detection, regression and cla

researcharxiv-cs-lg
21 Apr 2026
Research
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Beam-Plasma Collective Oscillations in Intense Charged-Particle Beams: Dielectric Response Theory, Langmuir Wave Dispersion, and Unsupervised Detection via Prometheus

DGX agent

arXiv:2603.10457v3 Announce Type: replace-cross Abstract: We develop a theoretical and computational framework for beam-plasma collective oscillations in intense charged-particle beams at intermediate

researcharxiv-cs-lg
21 Apr 2026
Research

Before You Interpret the Profile: Validity Scaling for LLM Metacognitive Self-Report

DGX agent

arXiv:2604.17707v1 Announce Type: new Abstract: Clinical personality assessment screens response validity before interpreting substantive scales. LLM evaluation does not. We apply the validity scaling

researcharxiv-cs-cl
21 Apr 2026
Model Releases

BEFT: Bias-Efficient Fine-Tuning of Language Models in Low-Data Regimes

DGX agent

arXiv:2509.15974v2 Announce Type: replace Abstract: Fine-tuning the bias terms of large language models (LLMs) has the potential to achieve unprecedented parameter efficiency while maintaining competi

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

BenchMarker: An Education-Inspired Toolkit for Highlighting Flaws in Multiple-Choice Benchmarks

DGX agent

arXiv:2602.06221v2 Announce Type: replace Abstract: Multiple-choice question answering (MCQA) is standard in NLP, but benchmarks lack rigorous quality control. We present BenchMarker, an education-ins

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Benchmarking Real-Time Question Answering via Executable Code Workflows

DGX agent

arXiv:2604.16349v1 Announce Type: cross Abstract: Retrieving real-time information is a fundamental capability for search-integrated agents in real-world applications. However, existing benchmarks are

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Benchmarking System Dynamics AI Assistants: Cloud Versus Local LLMs on CLD Extraction and Discussion

DGX agent

arXiv:2604.18566v1 Announce Type: cross Abstract: We present a systematic evaluation of large language model families -- spanning both proprietary cloud APIs and locally-hosted open-source models -- o

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

BengaliMoralBench: A Benchmark for Auditing Moral Reasoning in Large Language Models within Bengali Language and Culture

DGX agent

arXiv:2511.03180v2 Announce Type: replace Abstract: As multilingual Large Language Models (LLMs) gain traction across South Asia, their alignment with local ethical norms, particularly for Bengali, sp

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

Better with Less: Tackling Heterogeneous Multi-Modal Image Joint Pretraining via Conditioned and Degraded Masked Autoencoder

DGX agent

arXiv:2604.16952v1 Announce Type: new Abstract: Learning robust representations across extremely heterogeneous modalities remains a fundamental challenge in multi-modal vision. As a critical and profo

safetyarxiv-cs-cv
21 Apr 2026
Tutorials

Beyond Attack Success Rate: A Multi-Metric Evaluation of Adversarial Transferability in Medical Imaging Models

DGX agent

arXiv:2604.16532v1 Announce Type: new Abstract: While deep learning systems are becoming increasingly prevalent in medical image analysis, their vulnerabilities to adversarial perturbations raise seri

tutorialsarxiv-cs-cv
21 Apr 2026
Research

Beyond Binary Contrast: Modeling Continuous Skeleton Action Spaces with Transitional Anchors

DGX agent

arXiv:2604.17914v1 Announce Type: new Abstract: Self-supervised contrastive learning has emerged as a powerful paradigm for skeleton-based action recognition by enforcing consistency in the embedding

researcharxiv-cs-cv
21 Apr 2026
Research

Beyond Black-Box Labels: Interpretable Criteria for Diagnosing SubjectiveNLP Tasks

DGX agent

arXiv:2604.17022v1 Announce Type: new Abstract: Subjective NLP datasets typically aggregate annotator judgments into a single gold label, making it difficult to diagnose whether disagreement reflects

researcharxiv-cs-cl
21 Apr 2026
Model Releases

Beyond Feature Fusion: Contextual Bayesian PEFT for Multimodal Uncertainty Estimation

DGX agent

arXiv:2604.16615v1 Announce Type: new Abstract: We introduce CoCo-LoRA, a multimodal, uncertainty-aware parameter-efficient fine-tuning method for text prediction tasks accompanied by audio context. E

model-releasesarxiv-cs-lg
21 Apr 2026
Research

Beyond Fine-Tuning: In-Context Learning and Chain-of-Thought for Reasoned Distractor Generation

DGX agent

arXiv:2604.17574v1 Announce Type: new Abstract: Distractor generation (DG) remains a labor-intensive task that still significantly depends on domain experts. The task focuses on generating plausible y

researcharxiv-cs-cl
21 Apr 2026
Model Releases

Beyond 'I Don't Know': Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty

DGX agent

arXiv:2604.17293v1 Announce Type: new Abstract: Reliable Large Language Models (LLMs) should abstain when confidence is insufficient. However, prior studies often treat refusal as a generic 'I don't k

model-releasesarxiv-cs-cl
21 Apr 2026
Local Ai

Beyond Memorization: Extending Reasoning Depth with Recurrence, Memory and Test-Time Compute Scaling

DGX agent

arXiv:2508.16745v2 Announce Type: replace Abstract: Reasoning is a core capability of large language models, yet how multi-step reasoning is learned and executed remains unclear. We study this questio

local-aiarxiv-cs-lg
21 Apr 2026
Safety

Beyond Overlap Metrics: Rewarding Reasoning and Preferences for Faithful Multi-Role Dialogue Summarization

DGX agent

arXiv:2604.17188v1 Announce Type: new Abstract: Multi-role dialogue summarization requires modeling complex interactions among multiple speakers while preserving role-specific information and factual

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

Beyond Pattern Matching: Seven Cross-Domain Techniques for Prompt Injection Detection

DGX agent

arXiv:2604.18248v1 Announce Type: cross Abstract: Current open-source prompt-injection detectors converge on two architectural choices: regular-expression pattern matching and fine-tuned transformer c

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Beyond Reproduction: A Paired-Task Framework for Assessing LLM Comprehension and Creativity in Literary Translation

DGX agent

arXiv:2604.18169v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for creative tasks such as literary translation. Yet translational creativity remains underexplored a

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Beyond Static Benchmarks: Synthesizing Harmful Content via Persona-based Simulation for Robust Evaluation

DGX agent

arXiv:2604.17020v1 Announce Type: new Abstract: Static benchmarks for harmful content detection face limitations in scalability and diversity, and may also be affected by contamination from web-scale

researcharxiv-cs-cl
21 Apr 2026
Research

Beyond the Failures: Rethinking Foundation Models in Pathology

DGX agent

arXiv:2510.23807v5 Announce Type: replace-cross Abstract: Despite their successes in vision and language, foundation models have stumbled in pathology, revealing low accuracy, instability, and heavy c

researcharxiv-cs-cv
21 Apr 2026
Research

Beyond URLs: Metadata Diversity and Position for Efficient LLM Pretraining

DGX agent

arXiv:2511.21613v2 Announce Type: replace Abstract: Incorporating metadata in Large Language Models (LLMs) pretraining has recently emerged as a promising approach to accelerate training. However prio

researcharxiv-cs-cl
21 Apr 2026
Research

Beyond Verifiable Rewards: Rubric-Based GRM for Reinforced Fine-Tuning SWE Agents

DGX agent

arXiv:2604.16335v1 Announce Type: new Abstract: Despite recent progress in Large Language Model (LLM) Agents for Software Engineering (SWE) tasks, end-to-end fine-tuning typically relies on verifiable

researcharxiv-cs-lg
21 Apr 2026
Model Releases

Beyond Word Boundaries: A Hebrew Coreference Benchmark and an Evaluation Protocol for Morphologically Complex Text

DGX agent

arXiv:2604.17108v1 Announce Type: new Abstract: Coreference Resolution (CR) is a fundamental NLP task critical for long-form tasks as information extraction, summarization, and many business applicati

model-releasesarxiv-cs-cl
21 Apr 2026
Research

BhashaSutra: A Task-Centric Unified Survey of Indian NLP Datasets, Corpora, and Resources

DGX agent

arXiv:2604.18423v1 Announce Type: new Abstract: India's linguistic landscape, spanning 22 scheduled languages and hundreds of marginalized dialects, has driven rapid growth in NLP datasets, benchmarks

researcharxiv-cs-cl
21 Apr 2026
Model Releases

Bi-LoRA: Efficient Sharpness-Aware Minimization for Fine-Tuning Large-Scale Models

DGX agent

arXiv:2508.19564v2 Announce Type: replace Abstract: Fine-tuning large-scale pre-trained models with limited data presents significant challenges for generalization. While Sharpness-Aware Minimization

model-releasesarxiv-cs-lg
21 Apr 2026
Safety

Bias-constrained multimodal intelligence for equitable and reliable clinical AI

DGX agent

arXiv:2604.16884v1 Announce Type: new Abstract: The integration of medical imaging and clinical text has enabled the emergence of generalist artificial intelligence (AI) systems for healthcare. Howeve

safetyarxiv-cs-cv
21 Apr 2026
Safety

BIASEDTALES-ML: A Multilingual Dataset for Analyzing Narrative Attribute Distributions in LLM-Generated Stories

DGX agent

arXiv:2604.17008v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used to generate narrative content, including children's stories, which play an important role in social a

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

Bielik Guard: Efficient Polish Language Safety Classifiers for LLM Content Moderation

DGX agent

arXiv:2602.07954v4 Announce Type: replace Abstract: As Large Language Models (LLMs) become increasingly deployed in Polish language applications, the need for efficient and accurate content safety cla

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Bilinear Input Modulation for Mamba: Koopman Bilinear Forms for Memory Retention and Multiplicative Computation

DGX agent

arXiv:2604.17221v1 Announce Type: cross Abstract: Selective State Space Models (SSMs), notably Mamba, employ diagonal state transitions that limit both memory retention and bilinear computational capa

researcharxiv-cs-lg
21 Apr 2026
Research

BioVLM: Routing Prompts, Not Parameters, for Cross-Modality Generalization in Biomedical VLMs

DGX agent

arXiv:2604.17629v1 Announce Type: new Abstract: Pretrained biomedical vision-language models (VLMs) such as BioMedCLIP perform well on average but often degrade on challenging modalities where inter-c

researcharxiv-cs-cv
21 Apr 2026
Hardware

Bit-Flip Vulnerability of Shared KV-Cache Blocks in LLM Serving Systems

DGX agent

arXiv:2604.17249v1 Announce Type: cross Abstract: Rowhammer on GPU DRAM has enabled adversarial bit flips in model weights; shared KV-cache blocks in LLM serving systems present an analogous but previ

hardwarearxiv-cs-lg
21 Apr 2026
Tutorials

Block-encodings as programming abstractions: The Eclipse Qrisp BlockEncoding Interface

DGX agent

arXiv:2604.18276v1 Announce Type: cross Abstract: Block-encoding is a foundational technique in modern quantum algorithms, enabling the implementation of non-unitary operations by embedding them into

tutorialsarxiv-cs-lg
21 Apr 2026
Agents

BOIL: Learning Environment Personalized Information

DGX agent

arXiv:2604.17137v1 Announce Type: new Abstract: Navigating complex environments poses challenges for multi-agent systems, requiring efficient extraction of insights from limited information. In this p

agentsarxiv-cs-lg
21 Apr 2026
Safety

Boltzmann Machine Learning with a Parallel, Persistent Markov chain Monte Carlo method for Estimating Evolutionary Fields and Couplings from a Protein Multiple Sequence Alignment

DGX agent

arXiv:2604.18022v1 Announce Type: cross Abstract: The inverse Potts problem for estimating evolutionary single-site fields and pairwise couplings in homologous protein sequences from their single-site

safetyarxiv-cs-lg
21 Apr 2026
Agents

Bolzano: Case Studies in LLM-Assisted Mathematical Research

DGX agent

arXiv:2604.16989v1 Announce Type: new Abstract: We report new results on six problems in mathematics and theoretical computer science, produced with the assistance of Bolzano, an open-source multi-age

agentsarxiv-cs-cl
21 Apr 2026
Model Releases

BOOKAGENT: Orchestrating Safety-Aware Visual Narratives via Multi-Agent Cognitive Calibration

DGX agent

arXiv:2604.16541v1 Announce Type: new Abstract: Recent advancements in Large Generative Models (LGMs) have revolutionized multi-modal generation. However, generating illustrated storybooks remains an

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

BOP-ASK: Object-Interaction Reasoning for Vision-Language Models

DGX agent

arXiv:2511.16857v3 Announce Type: replace Abstract: Vision Language Models (VLMs) have achieved impressive performance on spatial reasoning benchmarks, yet these evaluations mask critical weaknesses i

model-releasesarxiv-cs-cv
21 Apr 2026
Research

Bounded Graph Clustering with Graph Neural Networks

DGX agent

arXiv:2512.05623v2 Announce Type: replace Abstract: In community detection, many methods require the user to specify the number of clusters in advance since an exhaustive search over all possible valu

researcharxiv-cs-lg
21 Apr 2026
Safety

Bounded Ratio Reinforcement Learning

DGX agent

arXiv:2604.18578v1 Announce Type: new Abstract: Proximal Policy Optimization (PPO) has become the predominant algorithm for on-policy reinforcement learning due to its scalability and empirical robust

safetyarxiv-cs-lg
21 Apr 2026
Research

Brain-CLIPLM: Decoding Compressed Semantic Representations in EEG for Language Reconstruction

DGX agent

arXiv:2604.16370v1 Announce Type: new Abstract: Decoding natural language from non-invasive electroencephalography (EEG) remains fundamentally limited by low signal-to-noise ratio and restricted infor

researcharxiv-cs-cl
21 Apr 2026
Model Releases

Brain-Inspired Capture: Evidence-Driven Neuromimetic Perceptual Simulation for Visual Decoding

DGX agent

arXiv:2604.17927v1 Announce Type: new Abstract: Visual decoding of neurophysiological signals is a critical challenge for brain-computer interfaces (BCIs) and computational neuroscience. However, curr

model-releasesarxiv-cs-cv
21 Apr 2026
Agents

BrainMem: Brain-Inspired Evolving Memory for Embodied Agent Task Planning

DGX agent

arXiv:2604.16331v1 Announce Type: cross Abstract: Embodied task planning requires agents to execute long-horizon, goal-directed actions in complex 3D environments, where success depends on both immedi

agentsarxiv-cs-cv
21 Apr 2026
Applications

Bridge-Centered Metapath Classification Using R-GCN-VGAE for Disaster-Resilient Maintenance Decisions

DGX agent

arXiv:2604.18399v1 Announce Type: new Abstract: Daily infrastructure management in preparation for disasters is critical for urban resilience. When bridges remain resilient against disaster-induced ex

applicationsarxiv-cs-lg
21 Apr 2026
Safety

BRIDGE the Gap: Mitigating Bias Amplification in Automated Scoring of English Language Learners via Inter-group Data Augmentation

DGX agent

arXiv:2602.23580v2 Announce Type: replace Abstract: In the field of educational assessment, automated scoring systems increasingly rely on deep learning and large language models (LLMs). However, thes

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

BridgeEQA: Virtual Embodied Agents for Real Bridge Inspections

DGX agent

arXiv:2511.12676v2 Announce Type: replace Abstract: Deploying embodied agents that can answer questions about their surroundings in realistic real-world settings remains difficult, partly due to the s

model-releasesarxiv-cs-cv
21 Apr 2026
Research

Bridging Coarse and Fine Recognition: A Hybrid Approach for Open-Ended Multi-Granularity Object Recognition in Interactive Educational Games

DGX agent

arXiv:2604.16785v1 Announce Type: new Abstract: Recent advances in Multimodal Large Language Models (MLLMs) have enabled open-ended object recognition, yet they struggle with fine-grained tasks. In co

researcharxiv-cs-cv
21 Apr 2026
Local Ai

Bridging the Culture Gap: A Framework for LLM-Driven Socio-Cultural Localization of Math Word Problems in Low-Resource Languages

DGX agent

arXiv:2508.14913v4 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated significant capabilities in solving mathematical problems expressed in natural language. However, mul

local-aiarxiv-cs-cl
21 Apr 2026
← Previous
1…11031104110511061107…1247
Next →