AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
Human
87,617Total entries
1Added by human
87,616Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,383 results
26 May 2026

Faithfulness as Information Flow: Evaluating and Training Faithful Chain-of-Thought Reasoning

SafetyDGX agent

arXiv:2605.24286v1 Announce Type: cross Abstract: Chain-of-thought (CoT) reasoning is useful for monitoring language models only when the reasoning trace faithfully reflects the computation that produ

Faithfulness Metrics Don't Measure Faithfulness: A Meta-Evaluation with Ground Truth

Model ReleasesDGX agent

arXiv:2605.25052v1 Announce Type: new Abstract: Chains of thought (CoTs) have become central in interpreting and auditing behaviors of large language models. Yet growing evidence suggests that these t

False Fixed Points: Kantian Feedback, Stable Miscalibration, and Representational Compression in LLMs

Local AiDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2510.14925v4 Announce Type: replace Abstract: High-confidence errors in large language models are often treated as fragile failures. We study an alternative: some errors may be false fixed point

Feature Learning in Wide Neural Networks under muP: Identifiability and Sparse-Dictionary Decomposition of the Mean-Field Limit

Model ReleasesDGX agent

arXiv:2605.24710v1 Announce Type: new Abstract: We establish four structural results for feature learning in wide two-layer neural networks under the Maximal Update Parametrization (muP). First, we pr

Feature Lottery? A Bifurcation Theory of Concept Emergence

ResearchDGX agent

arXiv:2605.24057v1 Announce Type: cross Abstract: Neural networks acquire structured representations at specific moments during training, yet identifying these transitions typically relies on retrospe

Feature Resemblance: Towards a Theoretical Understanding of Analogical Reasoning in Transformers

ResearchDGX agent

arXiv:2603.05143v3 Announce Type: replace Abstract: Understanding reasoning in large language models is complicated by evaluations that conflate multiple reasoning types. We isolate analogical reasoni

Federated Learning over Human-Body Communication for On-Body Edge Intelligence: A Survey, Taxonomy, and BODYFED-HBC Scheduling Vignette

Local AiDGX agent

arXiv:2605.24062v1 Announce Type: cross Abstract: Human-body communication (HBC) is a promising physical substrate for wearable body-area networks because it can localize communication around the body

Federated Sketching LoRA: A Flexible Framework for Heterogeneous Collaborative Fine-Tuning of LLMs

ResearchDGX agent

arXiv:2501.19389v4 Announce Type: replace Abstract: Fine-tuning large language models (LLMs) on resource-constrained clients remains a challenging problem. Recent works have fused low-rank adaptation

Fermi-Dirac machines as quantizations of neurons

TutorialsDGX agent

arXiv:2605.24386v1 Announce Type: cross Abstract: Fermi-Dirac machines were proposed recently as an approach to solving semidefinite optimization problems on quantum computers. Here, we reinterpret th

Few-Shot Neural Differentiable Simulator: Real-to-Sim Rigid-Contact Modeling

SafetyDGX agent

arXiv:2603.06218v2 Announce Type: replace Abstract: Accurate physics simulation is essential for robotic learning and control, yet analytical simulators often fail to capture complex contact dynamics,

FG-CLIP 2: A Bilingual Fine-grained Vision-Language Alignment Model

Model ReleasesDGX agent

arXiv:2510.10921v3 Announce Type: replace-cross Abstract: Fine-grained vision-language understanding requires precise alignment between visual content and linguistic descriptions, a capability that re

Filtered Posterior Mean Collections: A Unified Framework for Analytical Models of Diffusion Generalization

ResearchDGX agent

arXiv:2605.24192v1 Announce Type: cross Abstract: The neural-network denoising functions which form the backbone of image diffusion models are remarkably consistent in their generalization behaviour a

Fine-Tuning and Serving Gemma 4 31B on Google Cloud TPU: A Technical Comparison with GPU Baselines

Model ReleasesDGX agent

arXiv:2605.25645v1 Announce Type: cross Abstract: We present the first end-to-end demonstration of fine-tuning and serving Google's Gemma 4 31B model on TPU hardware, providing an empirical comparison

Fine-Tuning Language Models to Know What They Know

Model ReleasesDGX agent

arXiv:2602.02605v2 Announce Type: replace-cross Abstract: Evaluating true metacognition in Large Language Models (LLMs) is difficult due to biases and heuristics. This paper presents a framework to me

Fine-Tuning Masked Diffusion for Provable Self-Correction

ResearchDGX agent

arXiv:2510.01384v4 Announce Type: replace Abstract: A natural desideratum for generative models is self-correction--detecting and revising low-quality tokens at inference. While Masked Diffusion Model

Fine-Tuning Over Architectural Complexity: Broad-Coverage PII Detection on PIIBench with DeBERTa

ResearchDGX agent

arXiv:2605.25816v1 Announce Type: cross Abstract: Personally identifiable information (PII) detection systems are frequently trained within narrow source or domain boundaries, limiting coverage when d

First, do no harm: Breaking suicidogenic echo chambers in media recommendation

SafetyDGX agent

arXiv:2605.25258v1 Announce Type: cross Abstract: Recommender systems generally optimises user engagement, but this approach is dangerous in mental health contexts. When vulnerable users show signs of

Flat Minima and Generalization: Insights from Stochastic Convex Optimization

SafetyDGX agent

arXiv:2511.03548v2 Announce Type: replace Abstract: Understanding the generalization behavior of learning algorithms is a central goal of learning theory. A recently emerging explanation is that learn

FLINGO -- Instilling ASP Expressiveness into Linear Integer Constraints

ApplicationsDGX agent

arXiv:2602.09620v2 Announce Type: replace Abstract: Constraint Answer Set Programming (CASP) is a hybrid paradigm that enriches Answer Set Programming (ASP) with numerical constraint processing, a cru

FLOATBench: A Dataset and Benchmark for Floating Offshore Wind Turbine Tower Fatigue

Model ReleasesDGX agent

arXiv:2605.25717v1 Announce Type: new Abstract: Most of the world's offshore wind resource lies in waters too deep for fixed-bottom foundations, making floating offshore wind turbines (FOWTs) essentia

FloorplanQA: A Benchmark for Spatial Reasoning in LLMs using Structured Representations

Model ReleasesDGX agent

arXiv:2507.07644v4 Announce Type: replace Abstract: We introduce FloorplanQA, a diagnostic benchmark for evaluating spatial reasoning in large language models (LLMs). FloorplanQA is grounded in struct

FLoRIST: Singular Value Thresholding for Efficient and Accurate Federated Fine-Tuning of Large Language Models

Model ReleasesDGX agent

arXiv:2506.09199v2 Announce Type: replace-cross Abstract: Integrating Low-Rank Adaptation (LoRA) into federated learning offers a promising solution for parameter-efficient fine-tuning of Large Langua

FoodMonitor: Benchmarking MLLMs for Explainable Compliance Analysis

Model ReleasesDGX agent

arXiv:2605.24503v1 Announce Type: cross Abstract: As AI-powered compliance monitoring becomes increasingly important in public governance and industrial safety, the ability to provide verifiable evide

Forgettable Federated Linear Learning with Certified Data Unlearning

ResearchDGX agent

arXiv:2306.02216v3 Announce Type: replace Abstract: Federated Learning (FL) enables collaborative model training across distributed clients while preserving user privacy. Recently, Federated Unlearnin

Forgetting in Language Models: Capacity, Optimization, and Self-Generated Replay

ResearchDGX agent

arXiv:2605.26097v1 Announce Type: new Abstract: Models trained on a new task typically degrade on prior tasks, a phenomenon known as forgetting. Traditionally, mitigating forgetting has required repla

Forgotten Words: Benchmarking NeoBERT for Dementia Detection in Low-Resource Conversational Filipino and English Speech

ResearchDGX agent

arXiv:2605.26007v1 Announce Type: new Abstract: Dementia detection from spontaneous speech offers a scalable approach to cognitive screening, yet NLP systems remain predominantly English-centric. This

Found in Conversation: LLMs Teach Themselves to Close the Multi-Turn Gap

Model ReleasesDGX agent

arXiv:2605.24432v1 Announce Type: new Abstract: Large Language Model (LLM) interactions are typically underspecified, with users clarifying all necessary details across multiple conversational turns.

FOUND-IT: Foundation-model-first Task-driven 3D Scene Graphs with Granularity on Demand

Model ReleasesDGX agent

arXiv:2605.25371v1 Announce Type: new Abstract: We present the first approach to build hierarchical task-driven 3D scene graphs of arbitrary indoor or outdoor environments using an uncalibrated monocu

Fourier Feature Pyramids for Physics-Informed Neural Networks

Model ReleasesDGX agent

arXiv:2605.24278v1 Announce Type: new Abstract: We present an improved neural field architecture for solving partial differential equations (PDEs). Current physics-informed neural networks (PINNs) pro

FragmentNet: Adaptive Graph Fragmentation for Graph-to-Sequence Molecular Representation Learning

ResearchDGX agent

arXiv:2502.01184v2 Announce Type: replace-cross Abstract: Molecular representation learning methods typically tokenize molecules as individual atoms or use rigid, rule-based fragment decompositions, l

Fredholm Neural Networks for inverse problems in elliptic PDEs

ResearchDGX agent

arXiv:2507.06038v4 Announce Type: replace-cross Abstract: Building on our previous work on Fredholm Neural Networks (Fredholm NNs/ FNNs) for solving integral equations, we extend the framework to inve

Frequency Matters: Fast Model-Agnostic Data Curation for Pruning and Quantization

ResearchDGX agent

arXiv:2603.16105v3 Announce Type: replace-cross Abstract: Post-training model compression is essential for enhancing the portability of Large Language Models (LLMs) while preserving their performance.

From Accounting to Coordination: A Virtual Water-Aware Electricity-Computation-Water Nexus Framework for Data Center Dispatch

ResearchDGX agent

arXiv:2605.25854v1 Announce Type: new Abstract: The expansion of data centers (DCs) drives a sustained increase in electricity demand and associated water withdrawals at generation sites. These withdr

From Accuracy to Auditability: A Survey of Determinism in Financial AI Systems

AgentsDGX agent

arXiv:2605.23955v1 Announce Type: new Abstract: Deploying machine learning in regulated financial environments -- credit risk, fraud detection, and anti-money laundering -- exposes critical vulnerabil

From Automation to Collaboration: Human-in-the-Loop Methods for Safe and Trustworthy NLP

SafetyDGX agent

arXiv:2605.25226v1 Announce Type: new Abstract: Large language models are widely deployed in high-stakes NLP tasks, yet risks such as bias, hallucination, adversarial vulnerability and unreliable gene

From DPPs to k-DPPs: identifiability analysis via spectral decomposition

Model ReleasesDGX agent

arXiv:2605.25526v1 Announce Type: cross Abstract: We study the geometry of determinantal point processes (DPPs) through the spectral decomposition L=ULambda U^{op}. The spectrum Lambda governs the car

From Facts to Insights: A Persona-Driven Dual Memory Framework and Dataset for Role-Playing Agents

Model ReleasesDGX agent

arXiv:2605.25693v1 Announce Type: new Abstract: While role-playing agents excel in short-term interactions, long-term conversations overwhelm context windows, motivating external memory frameworks. Cu

From Index to Equity: Pre-Training Transformers for Stock Return Prediction

Model ReleasesDGX agent

arXiv:2605.23962v1 Announce Type: cross Abstract: This research aims to leverage machine learning to improve stock price prediction and support informed investment decisions related to buying, selling

From Knowledge to Inference: Formalizing Specialized Public Health Reasoning on GlobalHealthAtlas

SafetyDGX agent

arXiv:2602.00491v2 Announce Type: replace Abstract: Public health reasoning requires population level inference grounded in scientific evidence, expert consensus, and safety constraints. However, it r

From Latent Space to Training Data: Explainable Specialization in Minimal MLPs

ResearchDGX agent

arXiv:2605.25939v1 Announce Type: cross Abstract: We here study whether training biases can make hidden neurons specialize in minimal one-hidden-layer MLPs, and whether such specialization improves pr

From Model Scaling to System Scaling: Scaling the Harness in Agentic AI

Model ReleasesDGX agent

arXiv:2605.26112v1 Announce Type: new Abstract: This paper studies the next major bottleneck in agentic AI as system scaling, not only model scaling: the design of auditable, persistent, modular, and

From Multi-Agent Systems and the Semantic Web to Agentic AI: A Unified Narrative of the Web of Agents

SafetyDGX agent

arXiv:2507.10644v4 Announce Type: replace Abstract: The Web of Agents (WoA) transforms the document-centric Web into an environment of autonomous agents acting on users' behalf, a vision newly tractab

From One-Pass SGD to Data Reuse: Mini-Batch Scaling Laws in Sketched Linear Regression

Model ReleasesDGX agent

arXiv:2605.24316v1 Announce Type: new Abstract: Scaling laws provide compact descriptions of how prediction error varies with compute, model size, and data, but existing theory mainly treats single-sa

From Prompt Optimization to Multi-Dimensional Credibility Evaluation: Enhancing Trustworthiness of Chinese LLM-Generated Liver MRI Reports -- with Preliminary Extension to Lung Cancer

Model ReleasesDGX agent

arXiv:2510.23008v3 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated promising performance in generating diagnostic conclusions from imaging findings, thereby supporting

From Reasoning to Code: GRPO Optimization for Underrepresented Languages

SafetyDGX agent

arXiv:2506.11027v3 Announce Type: replace-cross Abstract: Generating accurate and executable code using Large Language Models (LLMs) remains a significant challenge for underrepresented programming la

From Simulation to Enaction: Post-trained language models recognize and react to their own generations

SafetyDGX agent

arXiv:2605.25459v1 Announce Type: cross Abstract: Language models are pretrained as passive predictors with no incentive to model the consequences of their own outputs. Post-training changes this: a m

From Sleep Staging to Spindle Detection: A Case Study on End-to-End Automated Sleep Analysis

ApplicationsDGX agent

arXiv:2505.05371v2 Announce Type: replace-cross Abstract: Automation of sleep analysis, including both macrostructural (sleep stages) and microstructural (e.g., sleep spindles) elements, promises to e

From Theory to Decision Rule: Calibrating the Noisy-Label Crossover for Vision-Language Model Weak Supervision Across Three Medical-Imaging Benchmarks

Model ReleasesDGX agent

arXiv:2605.24771v1 Announce Type: cross Abstract: Classical noisy-label theory predicts that downstream performance under weak supervision is bounded above by the labeler's accuracy, implying a sharp

FrontierOR: Benchmarking LLMs' Capacity for Efficient Algorithm Design in Large-Scale Optimization

Model ReleasesDGX agent

arXiv:2605.25246v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for optimization modeling and solver-code generation, yet practical operations research and optimizat

Full-Spectrum Graph Neural Networks: Expressive and Scalable

ResearchDGX agent

arXiv:2605.05759v2 Announce Type: replace Abstract: It is well established that spectral graph neural networks (GNNs) can universally approximate node signals; however, their expressive power remains

Fundamental Limitation in Explaining AI

ResearchDGX agent

arXiv:2605.24727v1 Announce Type: new Abstract: While large-scale models such as LLMs and diffusion models have achieved practical success, public institutions have emphasized the importance of explai

Fundamental Limits for Sensor-Based Control via the Gibbs Variational Principle

ResearchDGX agent

arXiv:2603.18454v2 Announce Type: replace-cross Abstract: Fundamental limits on the performance of feedback controllers are essential for benchmarking algorithms, guiding sensor selection, and certify

Fusion Embedding for Pose-Guided Person Image Synthesis with Diffusion Model

Model ReleasesDGX agent

arXiv:2412.07333v2 Announce Type: replace-cross Abstract: Pose-Guided Person Image Synthesis (PGPIS) aims to generate human images in specified poses while preserving the identity and appearance of a

FusionCore: A 23-State Unscented Kalman Filter for IMU, Wheel Encoder, GPS, and Visual SLAM Fusion in ROS 2

SafetyDGX agent

arXiv:2605.25239v1 Announce Type: new Abstract: We present FusionCore, an open-source ROS 2 sensor fusion package that fuses IMU, wheel encoder odometry, GPS, and Visual SLAM pose into a single 100 Hz

Future-KL Regularized GRPO: Process-Level Credit Assignment from f-Divergence Regularization

Local AiDGX agent

arXiv:2601.10201v2 Announce Type: replace-cross Abstract: Group Relative Policy Optimization (GRPO) is widely used for critic-free Large Language Model (LLM) post-training, but its KL regularization i

FuXi-Nowcast: Environment-conditioned deep learning for severe convection nowcasting

ResearchDGX agent

arXiv:2512.08974v2 Announce Type: replace-cross Abstract: Severe convection produces localized hazards that often require warnings before radar echoes fully reveal storm development. Convective initia

Fuzzy, Neutrosophic, and Uncertain Graph Theory: Properties and Applications

ResearchDGX agent

arXiv:2605.23936v1 Announce Type: new Abstract: This book presents a comprehensive and systematic survey of graph theory under uncertainty, with particular emphasis on the unifying role of the uncerta

Fuzzy PyTorch: Rapid Numerical Variability Evaluation for Deep Learning Models

ResearchDGX agent

arXiv:2605.25991v1 Announce Type: new Abstract: We introduce Fuzzy PyTorch, a framework for rapid evaluation of numerical variability in deep learning (DL) models. As DL is increasingly applied to div

G-DRAGON: Geospatial Reasoning and Dynamic Planning for Retrieval-Augmented Outdoor Navigation

Local AiDGX agent

arXiv:2605.25646v1 Announce Type: new Abstract: Autonomous ground robots operating in large-scale outdoor environments require both robust long-range navigation and fine-grained ''last-mile'' explorat

Game-Theoretic Modeling of Heterogeneous Investor Interactions for Stock Price Forecasting

Model ReleasesDGX agent

arXiv:2605.23953v1 Announce Type: cross Abstract: Accurate stock price forecasting has consistently remained a pivotal yet challenging FinTech task that underpins quantitative trading and investment d

← Previous
1…595596597598599…1040
Next →