AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
86,428Total entries
1Added by human
86,427Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
61,498 results
Model Releases

Behavior Cue Reasoning: Monitorable Reasoning Improves Efficiency and Safety through Oversight

DGX agent

arXiv:2605.07021v1 Announce Type: new Abstract: Reasoning in Large Language Models (LLMs) poses a challenge for oversight as many misaligned behaviors do not surface until reasoning concludes. To addr

model-releasesarxiv-cs-ai
11 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

Belief Memory: Agent Memory Under Partial Observability

DGX agent

arXiv:2605.05583v2 Announce Type: replace Abstract: LLM agents that operate over long context depend on external memory to accumulate knowledge over time. However, existing methods typically store eac

agentsarxiv-cs-ai
11 May 2026
Safety

Bellman Calibration for V-Learning in Offline Reinforcement Learning

DGX agent

arXiv:2512.23694v2 Announce Type: replace-cross Abstract: Reliable long-horizon value prediction is difficult in offline reinforcement learning because fitted value methods combine bootstrapping, func

safetyarxiv-cs-lg
11 May 2026
Model Releases

Benchmarked Yet Not Measured -- Generative AI Should be Evaluated Against Real-World Utility

DGX agent

arXiv:2605.06856v1 Announce Type: cross Abstract: Generative AI systems achieve impressive performance on standard benchmarks yet fail to deliver real-world utility, a disconnect we identify across 28

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Benchmarking EngGPT2-16B-A3B against Comparable Italian and International Open-source LLMs

DGX agent

arXiv:2605.07731v1 Announce Type: cross Abstract: This report benchmarks the performance of ENGINEERING Ingegneria Informatica S.p.A.'s EngGPT2MoE-16B-A3B LLM, a 16B parameter Mixture of Experts (MoE)

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Benchmarking Foundation Models for Renal Lesion Stratification in CT

DGX agent

arXiv:2605.07749v1 Announce Type: new Abstract: The rapid proliferation of open-source medical foundation models (FMs) raises a practical question: how well do their pre-trained representations transf

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Benchmarking World-Model Learning with Environment-Level Queries

DGX agent

arXiv:2510.19788v4 Announce Type: replace Abstract: World models are central to building AI agents capable of flexible reasoning and planning. Yet current evaluations (i) test only properties measurab

model-releasesarxiv-cs-ai
11 May 2026
Safety

Better Protein Function Prediction by Modeling Survivorship Bias

DGX agent

arXiv:2605.06879v1 Announce Type: new Abstract: Protein sequence data from nature exhibits survivorship bias: we only observe data from those organisms that survive and reproduce, while non-functional

safetyarxiv-cs-lg
11 May 2026
Safety

Beyond Confidence: Rethinking Self-Assessments for Performance Prediction in LLMs

DGX agent

arXiv:2605.07806v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in settings where reliable self-assessment is critical. Assessing model reliability has evolved fro

safetyarxiv-cs-ai
11 May 2026
Research

Beyond Defenses: Manifold-Aligned Regularization for Intrinsic 3D Point Cloud Robustness

DGX agent

arXiv:2605.07590v1 Announce Type: new Abstract: Despite extensive progress in point cloud robustness, existing methods primarily improve performance through augmentation or defense mechanisms, while o

researcharxiv-cs-cv
11 May 2026
Research

Beyond Distribution Estimation: Simplex Anchored Structural Inference Towards Universal Semi-Supervised Learning

DGX agent

arXiv:2605.07557v1 Announce Type: new Abstract: Semi-supervised learning faces significant challenges in realistic scenarios where labeled data is scarce and unlabeled data follows unknown, arbitrary

researcharxiv-cs-lg
11 May 2026
Model Releases

Beyond Factor Aggregation: Gauge-Aware Low-Rank Server Representations for Federated LoRA

DGX agent

arXiv:2605.06733v1 Announce Type: cross Abstract: Federated LoRA enables parameter-efficient adaptation of large language models under decentralized data and limited client resources.However, directly

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Beyond Factual Accuracy: Evaluating Global Reasoning Integrity in RAG Systems with LogicScore

DGX agent

arXiv:2601.15050v4 Announce Type: replace Abstract: Current evaluation methods for Retrieval Augmented Generation (RAG) suffer from extit{factual myopia}: they relentlessly emphasize factual accuracy

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Beyond GSD-as-Token: Continuous Scale Conditioning for Remote Sensing VLMs

DGX agent

arXiv:2605.07562v1 Announce Type: new Abstract: Remote sensing vision-language models (RS-VLMs) face a fundamental mismatch with natural-image counterparts: the same geographic object exhibits radical

model-releasesarxiv-cs-cv
11 May 2026
Safety

Beyond 'I cannot fulfill this request': Alleviating Rigid Rejection in LLMs via Label Enhancement

DGX agent

arXiv:2605.07883v1 Announce Type: new Abstract: Large Language Models (LLMs) rely on safety alignment to obey safe requests while refusing harmful ones. However, traditional refusal mechanisms often l

safetyarxiv-cs-cl
11 May 2026
Model Releases

Beyond Linear Attention: Softmax Transformers Implement In-Context Reinforcement Learning

DGX agent

arXiv:2605.07333v1 Announce Type: new Abstract: In-context reinforcement learning (ICRL) studies agents that, after pretraining, adapt to new tasks by conditioning on additional context without parame

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Beyond LoRA vs. Full Fine-Tuning: Gradient-Guided Optimizer Routing for LLM Adaptation

DGX agent

arXiv:2605.07111v1 Announce Type: cross Abstract: Recent literature on fine-tuning Large Language Models highlights a fundamental debate. While Full Fine-Tuning (FFT) provides the representational pla

model-releasesarxiv-cs-ai
11 May 2026
Safety

Beyond Pairs: Your Language Model is Secretly Optimizing a Preference Graph

DGX agent

arXiv:2605.08037v1 Announce Type: cross Abstract: Direct Preference Optimization (DPO) aligns language models using pairwise preference comparisons, offering a simple and effective alternative to Rein

safetyarxiv-cs-ai
11 May 2026
Research

Beyond Reasoning: Reinforcement Learning Unlocks Parametric Knowledge in LLMs

DGX agent

arXiv:2605.07153v1 Announce Type: new Abstract: Reinforcement learning (RL) has achieved remarkable success in LLM reasoning, but whether it can also improve direct recall of parametric knowledge rema

researcharxiv-cs-cl
11 May 2026
Model Releases

Beyond Retrieval: A Multitask Benchmark and Model for Code Search

DGX agent

arXiv:2605.04615v2 Announce Type: replace-cross Abstract: Code search has usually been evaluated as first-stage retrieval, even though production systems rely on broader pipelines with reranking and d

model-releasesarxiv-cs-ai
11 May 2026
Applications

Beyond Single Ground Truth: Reference Monism as Epistemic Injustice in ASR Evaluation

DGX agent

arXiv:2605.07084v1 Announce Type: new Abstract: Automatic speech recognition (ASR) evaluation compares system output to ground truth transcripts, with Word Error Rate (WER) quantifying the distance be

applicationsarxiv-cs-cl
11 May 2026
Safety

Beyond State-Wise Mirror Descent: Offline Policy Optimization with Parametric Policies

DGX agent

arXiv:2602.23811v4 Announce Type: replace-cross Abstract: We investigate the theoretical aspects of offline reinforcement learning (RL) under general function approximation. While prior works (e.g., X

safetyarxiv-cs-ai
11 May 2026
Model Releases

Beyond the Black Box: Interpretability of Agentic AI Tool Use

DGX agent

arXiv:2605.06890v1 Announce Type: new Abstract: AI agents are promising for high-stakes enterprise workflows, but dependable deployment remains limited because tool-use failures are difficult to diagn

model-releasesarxiv-cs-ai
11 May 2026
Tutorials

Beyond the Wrapper: Identifying Artifact Reliance in Static Malware Classifiers using TRUSTEE

DGX agent

arXiv:2605.07034v1 Announce Type: cross Abstract: Modern cybersecurity relies heavily on static machine-learning-based malware classifiers. However, transformations such as packing and other non-seman

tutorialsarxiv-cs-lg
11 May 2026
Model Releases

BGM-IV: an AI-powered Bayesian generative modeling approach for instrumental variable analysis

DGX agent

arXiv:2605.07029v1 Announce Type: cross Abstract: Instrumental-variable (IV) regression enables causal estimation under endogeneity, but modern IV problems often involve nonlinear structural effects a

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Bi3: A Biplatform, Bicultural, Biperson Dataset for Social Robot Navigation

DGX agent

arXiv:2605.06863v1 Announce Type: new Abstract: We contribute Bi3, a dataset of social robot navigation among groups of people in a constrained lab space. Compared to prior data collection efforts for

model-releasesarxiv-cs-ro
11 May 2026
Safety

Bias and Uncertainty in LLM-as-a-Judge Estimation

DGX agent

arXiv:2605.06939v1 Announce Type: new Abstract: LLM-as-a-Judge evaluation has become a standard tool for assessing base model performance. However, characterizing performance via the naive estimator,

safetyarxiv-cs-lg
11 May 2026
Tutorials

Bifurcation Models: Learning Set-Valued Solution Maps with Weight-Tied Dynamics

DGX agent

arXiv:2605.07277v1 Announce Type: cross Abstract: Many scientific and combinatorial problems admit multiple correct solutions, not a single label. Standard supervised learning resolves this ambiguity

tutorialsarxiv-cs-ai
11 May 2026
Research

Bilevel Graph Structure Learning, Revisited: Inner-Channel Origins of the Reported Gain

DGX agent

arXiv:2605.07577v1 Announce Type: new Abstract: Bilevel graph structure learning is widely understood to improve graph neural networks by jointly optimizing model parameters and a learned graph struct

researcharxiv-cs-lg
11 May 2026
Model Releases

BioProVLA-Agent: An Affordable, Protocol-Driven, Vision-Enhanced VLA-Enabled Embodied Multi-Agent System with Closed-Loop-Capable Reasoning for Biological Laboratory Manipulation

DGX agent

arXiv:2605.07306v1 Announce Type: cross Abstract: Biological laboratory automation can reduce repetitive manual work and improve reproducibility, but reliable embodied execution in wet-lab environment

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

bispectrum: Selective G-Bispectra Made Practical

DGX agent

arXiv:2605.07270v1 Announce Type: new Abstract: Many machine learning tasks are invariant under the action of a group G of transformations: signal classification can be invariant under translations, i

model-releasesarxiv-cs-lg
11 May 2026
Research

Black-box model classification under the discriminative factorization

DGX agent

arXiv:2605.07878v1 Announce Type: new Abstract: Access to modern generative systems is often restricted to querying an API (the ``black-box' setting) and many properties of the system are unknown to t

researcharxiv-cs-lg
11 May 2026
Research

Bloom Filter Encoding for Machine Learning

DGX agent

arXiv:2512.19991v2 Announce Type: replace Abstract: We present a method that uses a Bloom filter transform to preprocess data for machine learning. Each sample is encoded into a compact bit-array repr

researcharxiv-cs-lg
11 May 2026
Agents

Bluetooth Phased-array Aided Inertial Navigation Using Factor Graphs: Experimental Verification

DGX agent

arXiv:2602.17407v2 Announce Type: replace-cross Abstract: Phased-array Bluetooth systems have emerged as a low-cost alternative for performing aided inertial navigation in GNSS-denied use cases such a

agentsarxiv-cs-ro
11 May 2026
Model Releases

BoHA: Blockwise Hadamard Product Adaptation for Parameter-Efficient Fine-Tuning

DGX agent

arXiv:2509.21637v2 Announce Type: replace Abstract: Parameter-efficient fine-tuning (PEFT) of large language models trains a small task-specific parameter set while keeping the pretrained model frozen

model-releasesarxiv-cs-lg
11 May 2026
Research

Bounded Fitting for Expressive Description Logics

DGX agent

arXiv:2605.07452v1 Announce Type: new Abstract: Bounded fitting is an attractive paradigm for learning logical formulas from labeled data examples that offers PAC-style generalization guarantees and c

researcharxiv-cs-ai
11 May 2026
Research

Breaking QAOA's Fixed Target Hamiltonian Barrier: A Fully Connected Quantum Boltzmann Machine via Bilevel Optimization

DGX agent

arXiv:2605.07473v1 Announce Type: cross Abstract: To overcome the limitations of classical partially connected Boltzmann machines and mainstream quantum Boltzmann machines (QBMs), this work extends th

researcharxiv-cs-lg
11 May 2026
Model Releases

Breaking Spatial Uniformity: Prior-Guided Mamba with Radial Serialization for Lens Flare Removal

DGX agent

arXiv:2605.07650v1 Announce Type: new Abstract: Lens flares, caused by complex optical aberrations, severely degrade image quality especially in nighttime photography. Although recent restoration meth

model-releasesarxiv-cs-cv
11 May 2026
Research

Breaking the Illusion: When Positive Meets Negative in Multimodal Decoding

DGX agent

arXiv:2605.06679v1 Announce Type: new Abstract: Vision-Language Models (VLMs) are frequently undermined by object hallucination, generating content that contradicts visual reality, due to an over-reli

researcharxiv-cs-lg
11 May 2026
Agents

BrickCraft: Visuomotor Skill Composition with Situated Manual Guidance for Long-Horizon Interlocking Brick Assembly

DGX agent

arXiv:2605.07605v1 Announce Type: new Abstract: Autonomous robotic assembly of interlocking bricks demands seamless integration of long-horizon task reasoning, spatial grounding, and fine-grained mani

agentsarxiv-cs-ro
11 May 2026
Model Releases

BRIDGE: Background Routing and Isolated Discrete Gating for Coarse-Mask Local Editing

DGX agent

arXiv:2605.07846v1 Announce Type: new Abstract: Coarse-mask local image editing asks a model to modify a user-indicated region while preserving the surrounding scene. In practice, however, rough masks

model-releasesarxiv-cs-cv
11 May 2026
Research

Bridging Textual Profiles and Latent User Embeddings for Personalization

DGX agent

arXiv:2605.06981v1 Announce Type: cross Abstract: Personalized systems rely on user representations to connect behavioral history with downstream recommendation applications. Existing methods typicall

researcharxiv-cs-cl
11 May 2026
Model Releases

Bridging the Last Mile of Circuit Design: PostEDA-Bench, a Hierarchical Benchmark for PPA Convergence and DRC Fixing

DGX agent

arXiv:2605.06936v1 Announce Type: cross Abstract: LLM-based agents are increasingly applied to the 'last mile' of Electronic Design Automation (EDA): repairing residual sign-off Design Rule Check (DRC

model-releasesarxiv-cs-ai
11 May 2026
Tutorials

Bringing Multimodal Large Language Models to Infrared-Visible Image Fusion Quality Assessment

DGX agent

arXiv:2605.06969v1 Announce Type: new Abstract: Infrared-Visible image fusion (IVIF) aims to integrate thermal information and detailed spatial structures into a single fused image to enhance percepti

tutorialsarxiv-cs-cv
11 May 2026
Model Releases

CA-SQL: Complexity-Aware Inference Time Reasoning for Text-to-SQL via Exploration and Compute Budget Allocation

DGX agent

arXiv:2605.08057v1 Announce Type: cross Abstract: While recent advancements in inference-time learning have improved LLM reasoning on Text-to-SQL tasks, current solutions still struggle to perform wel

model-releasesarxiv-cs-ai
11 May 2026
Safety

CalexNet: Soft Cascade-Aligned Training and Calibration for Lightweight Early-Exit Branches

DGX agent

arXiv:2509.08318v2 Announce Type: replace Abstract: Early-exit cascades over a frozen convolutional backbone enable adaptive inference but suffer from three sources of train-inference mismatch: branch

safetyarxiv-cs-cv
11 May 2026
Model Releases

Can Agents Price a Reaction? Evaluating LLMs on Chemical Cost Reasoning

DGX agent

arXiv:2605.07251v1 Announce Type: new Abstract: Large Language Models (LLMs) have become increasingly capable as tool-using agents, with benchmarks spanning diverse general agentic tasks. Yet rigorous

model-releasesarxiv-cs-ai
11 May 2026
Safety

Can David Beat Goliath? On Multi-Hop Reasoning with Resource-Constrained Agents

DGX agent

arXiv:2601.21699v2 Announce Type: replace Abstract: Multi-turn reasoning agents solve complex questions by decomposing them into intermediate retrieval or tool-use steps, for accumulating supporting e

safetyarxiv-cs-cl
11 May 2026
← Previous
1…950951952953954…1282
Next →