AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,037 results
2 Jun 2026

AnyEdit++: Adaptive Long-Form Knowledge Editing via Bayesian Surprise

ResearchDGX agent

arXiv:2606.01053v1 Announce Type: new Abstract: Editing complex, long-form knowledge in Large Language Models remains a significant challenge due to the difficulty of maintaining generation coherence.

Atmospheric Predictability Beyond 30 Days with Machine Learning

ResearchDGX agent

arXiv:2504.20238v2 Announce Type: replace-cross Abstract: Atmospheric predictability research has long held that rapid error growth at small spatial scales imposes an intrinsic limit of roughly two we

AutoIQ: An Ensemble Framework for Automatic Assessment of Geometric Distortion in Prostate Diffusion-Weighted Imaging

Local AiDGX agent

arXiv:2606.00393v1 Announce Type: cross Abstract: Geometric distortion in prostate diffusion-weighted imaging (DWI) can impair lesion localization and reduce the reliability of MRI-based clinical asse

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Automated Conjecture Resolution with Formal Verification

AgentsDGX agent

arXiv:2604.03789v2 Announce Type: replace-cross Abstract: Recent advances in large language models have significantly improved their ability to perform mathematical reasoning, extending from elementar

Automated Essay Scoring and Language Certification: Assessing Generalizability, Agreement and Validity for French

SafetyDGX agent

arXiv:2606.02009v1 Announce Type: new Abstract: In Automated Essay Scoring (AES), benchmarking practices have fostered minimalist evaluation practices, in contrast with the broader-view recommendation

b9468

Local AiDGX agent

B9468 is an intermediate build release of llama.cpp, the open-source C/C++ library that enables large language model inference on consumer hardware. This release continues the project's rapid developm

b9469

Local AiDGX agent

b9469 is an intermediate build release of llama.cpp, a C/C++ implementation that enables large language model inference on consumer hardware with minimal dependencies. Build releases like b9469 repres

b9473

Local AiDGX agent

B9473 is a release of llama.cpp, a C/C++ framework for large language model inference . The release represents a specific build commit from the ggml-org/llama.cpp repository. This entry likely documen

b9480

Local AiDGX agent

B9480 is a release of llama.cpp, an LLM inference project in C/C++ . The release likely contains updates, bug fixes, or feature improvements to the llama.cpp codebase for running large language models

b9483

Local AiDGX agent

Build b9483 is an intermediate release of llama.cpp, the C/C++ inference engine for running large language models locally. This build represents a specific commit snapshot from the ggml-org llama.cpp

Beyond Access: Guided LLM Scaffolding for Independent Learning in Undergraduate Statistics

SafetyDGX agent

arXiv:2606.01375v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly entering students' learning practices, but their educational value depends on whether they support reaso

Beyond Low-Rank: Low-Rank Sparse Prompting via Spiking Neural Network and Prompt Factorization

ResearchDGX agent

arXiv:2606.01945v1 Announce Type: new Abstract: Visual Prompting (VP) has emerged as an efficient paradigm for adapting large-scale pre-trained vision models to downstream tasks by incorporating learn

Beyond Procedure: Substantive Fairness in Conformal Prediction

SafetyDGX agent

arXiv:2602.16794v2 Announce Type: replace-cross Abstract: Conformal prediction (CP) offers distribution-free uncertainty quantification for machine learning models, yet its interplay with fairness in

Both Topology and Text Matter: Revisiting LLM-guided Out-of-Distribution Detection on Text-attributed Graphs

SafetyDGX agent

arXiv:2602.11641v2 Announce Type: replace Abstract: Text-attributed graphs (TAGs) associate nodes with textual attributes and graph structure, enabling GNNs to jointly model semantic and structural in

BranPO: Scalable Contrastive Branch Sampling for Long-Horizon Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2602.03719v2 Announce Type: replace Abstract: Agentic reinforcement learning enables large language models to perform multi-turn planning and tool use, but long-horizon training remains challeng

Bridging the Last Mile of Time Series Forecasting with LLM Agents

SafetyDGX agent

arXiv:2606.02497v1 Announce Type: new Abstract: Time series forecasting has advanced rapidly, especially with the emergence of foundation models that show strong zero-shot performance on numerical ext

C-LEAD: Contrastive Learning for Enhanced Adversarial Defense

TutorialsDGX agent

arXiv:2510.27249v2 Announce Type: replace Abstract: Deep neural networks (DNNs) have achieved remarkable success in computer vision tasks such as image classification, segmentation, and object detecti

CA-BED: Conversation-Aware Bayesian Experimental Design

ResearchDGX agent

arXiv:2606.01182v1 Announce Type: cross Abstract: Large Language Models (LLMs) excel at static reasoning tasks, yet their performance often degrades in interactive scenarios where information must be

CARES: Context-Aware Resolution Selector for VLMs

ResearchDGX agent

arXiv:2510.19496v3 Announce Type: replace-cross Abstract: Large vision-language models (VLMs) commonly process images at native or high resolution to remain effective across tasks. This inflates visua

Characterizing the Effect of Noise in Language Generation in the Limit

ResearchDGX agent

arXiv:2601.21237v2 Announce Type: replace-cross Abstract: Kleinberg and Mullainathan recently proposed a formal framework for studying the phenomenon of language generation, called language generation

Chunking Methods on Retrieval-Augmented Generation - Effectiveness Evaluation Against Computational Cost and Limitations

ResearchDGX agent

arXiv:2606.00881v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has demonstrated significant capabilities in enhancing the performance of Large Language Models (LLMs). One of the

Code2Math: Can Your Code Agent Effectively Evolve Math Problems Through Exploration?

AgentsDGX agent

arXiv:2603.03202v3 Announce Type: replace Abstract: As large language models (LLMs) advance their mathematical capabilities toward the IMO and research level, the scarcity of challenging, high-quality

Context-aware child-directed speech detection from long-form recordings

ResearchDGX agent

arXiv:2606.01134v1 Announce Type: cross Abstract: Automatically distinguishing child-directed speech from adult-directed speech in long-form recordings is key to scalable analyses of children's langua

Contrastive meta-domain adaptation for robust skin lesion classification across clinical and acquisition conditions

ResearchDGX agent

arXiv:2602.19857v2 Announce Type: replace Abstract: Deep learning models for dermatological image analysis remain sensitive to acquisition variability and domain-specific visual characteristics, leadi

CORE-MTL: Rethinking Gradient Balancing via Causal Orthogonal Representations

ResearchDGX agent

arXiv:2606.02221v1 Announce Type: new Abstract: Multi-task learning (MTL) aims to construct a joint model for multiple tasks by sharing a common representation across domains. To achieve this goal, ex

Counterfactual Intervention Feature Transfer for Visible-Infrared Person Re-identification

ResearchDGX agent

arXiv:2208.00967v4 Announce Type: replace Abstract: Graph-based models have achieved great success in person re-identification tasks recently, which compute the graph topology structure (affinities) a

Deep Learning as the Disciplined Construction of Tame Objects

ResearchDGX agent

arXiv:2509.18025v2 Announce Type: replace-cross Abstract: One can see deep-learning models as compositions of functions within the so-called tame geometry. In this expository note, we give an overview

Defenses & Enablers For Skill Injection Attacks on Terminal Based Agents

AgentsDGX agent

arXiv:2606.01567v1 Announce Type: cross Abstract: Large language model (LLM) agents increasingly rely on reusable skills i.e. documents describing task-specific procedures. However, this introduces a

Demystifying Multi-Agent Debate: The Role of Confidence and Diversity

AgentsDGX agent

arXiv:2601.19921v2 Announce Type: replace-cross Abstract: Multi-agent debate (MAD) is widely used to improve large language model (LLM) performance through test-time scaling, yet recent work shows tha

DerMAE: Improving skin lesion classification through conditioned latent diffusion and MAE distillation

Local AiDGX agent

arXiv:2602.19848v2 Announce Type: replace Abstract: Skin lesion classification datasets often suffer from severe class imbalance, with malignant cases significantly underrepresented, leading to biased

Dimension Reduction via Sum-of-Squares and Improved Clustering Algorithms for Non-Spherical Mixtures

ResearchDGX agent

arXiv:2411.12438v2 Announce Type: replace-cross Abstract: We develop a new approach for clustering non-spherical (i.e., arbitrary component covariances) Gaussian mixture models via a subroutine, based

Dive into Ambiguity: A*-Inspired Multi-Agents Commonsense Obfuscation Attack on LLM Prompts

SafetyDGX agent

arXiv:2606.01441v1 Announce Type: new Abstract: Large language models (LLMs) excel in reasoning and knowledge-intensive tasks but remain vulnerable to prompt-level adversarial attacks that preserve in

Dive into Waves: Morlet Spectral Transformer for Cross-Subject Emotion Decoding from EEG

ResearchDGX agent

arXiv:2606.00884v1 Announce Type: cross Abstract: We study cross-subject emotion recognition from EEG, a practically important yet challenging problem in brain-computer interfaces. Unlike tasks with c

DyLLM: Efficient Diffusion LLM Inference via Saliency-based Token Selection and Partial Attention

ResearchDGX agent

arXiv:2603.08026v2 Announce Type: replace-cross Abstract: Masked diffusion language models enable parallel token decoding, providing a promising alternative to the sequential nature of autoregressive

Dynamic Coordination Strategy Selection for Enterprise Multi-Agent Systems

SafetyDGX agent

arXiv:2606.00804v1 Announce Type: cross Abstract: Enterprise multi-agent systems increasingly expose multiple coordination patterns, but deployments often lack evidence for when to use consensus, deba

Dynamic Trust-Aware Sparse Communication Topology for LLM-Based Multi-Agent Consensus

AgentsDGX agent

arXiv:2606.01828v1 Announce Type: cross Abstract: Large language model-driven multi-agent systems enhance the reliability of complex reasoning tasks through multi-round deliberation, role specializati

Edge-aware Decoding for Neural Asymmetric Routing

ResearchDGX agent

arXiv:2606.02136v1 Announce Type: new Abstract: Neural asymmetric routing models increasingly encode directionality through matrix representations and asymmetry-aware attention. The final routing acti

Edge Prediction for Roof Wireframe Reconstruction with Transformers

ResearchDGX agent

arXiv:2606.02406v1 Announce Type: new Abstract: This paper presents a competitive solution to the S23DR Challenge 2026, which aims to reconstruct 3D house roof wireframe models from sparse SfM point c

Effects of Varying LLM Access on Essay Writing Behavior

ResearchDGX agent

arXiv:2606.00250v1 Announce Type: cross Abstract: Investigating the degree to which large language models (LLMs) affect teaching and learning in universities can help identify strategies for integrati

Efficient Synthetic Network Generation via Latent Embedding Reconstruction

ResearchDGX agent

arXiv:2606.00934v1 Announce Type: cross Abstract: Network data are ubiquitous across the social sciences, biology, and information systems. Generating realistic synthetic network data has broad applic

ETC: Extreme Token Compression via Task-aware Visual Information Distillation in VLMs

ResearchDGX agent

arXiv:2606.00543v1 Announce Type: new Abstract: In Vision-Language Models (VLMs), high-resolution images produce a large number of visual tokens, resulting in high computational costs and KV-cache ove

Evaluating Bivariate Causal Statements Based on Mutual Compatibility

ApplicationsDGX agent

arXiv:2606.00278v1 Announce Type: new Abstract: For many real-world systems, causal ground truth is difficult to obtain, making claims about causal effects hard to assess. We develop methods for evalu

Evidence-Gated LLM Priors for Multi-Objective Bayesian Optimization

TutorialsDGX agent

arXiv:2606.01730v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as heuristic advisors for black-box optimization, yet their suggestions and self-reported confidence

Family Matters: A Systematic Study of Spatial vs. Frequency Masking for Continual Test-Time Adaptation

SafetyDGX agent

arXiv:2512.08048v3 Announce Type: replace Abstract: Recent continual test-time adaptation (CTTA) methods adopt masked image modeling to stabilize learning under distribution shift, yet each treats its

Fast and Lightweight Novel View Synthesis with Differentiable Multiplane Image

ResearchDGX agent

arXiv:2606.02068v1 Announce Type: cross Abstract: Recently, novel view synthesis has witnessed remarkable progress, with mainstream methods such as Neural Radiance Fields (NeRF) and 3D Gaussian Splatt

Finding What Matters: Anchoring Context Knowledge with Evolving Indices for Iterative Retrieval

TutorialsDGX agent

arXiv:2601.16462v2 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) has become a dominant paradigm for mitigating hallucinations in Large Language Models (LLMs) by incorporating e

Frequentist Consistency of Prior-Data Fitted Networks for Causal Inference

SafetyDGX agent

arXiv:2603.12037v2 Announce Type: replace Abstract: Foundation models based on prior-data fitted networks (PFNs) have shown strong empirical performance in causal inference by framing the task as an i

From Global to Local: Learning Context-Aware Graph Representations for Document Classification and Summarization

ResearchDGX agent

arXiv:2603.00021v2 Announce Type: replace Abstract: Recent NLP systems commonly represent documents as linear token sequences. Although this captures sequential order, it can hinder modeling long-rang

From Layers to Submodules: Rethinking Granularity in Replacement-Based LLM Compression

ResearchDGX agent

arXiv:2606.02559v1 Announce Type: cross Abstract: Post-training compression of Large Language Models (LLMs) removes entire architectural components, either deleting them or replacing them with fitted

Graph-Augmented Retrieval for Cross-Entity Financial Sentiment Analysis: A Comparative Study

ApplicationsDGX agent

arXiv:2606.00062v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has become foundational for grounding large language models in domain-specific corpora, yet conventional vector-bas

Graph Transfer Learning via Shared Latent Geometry: Theory and Applications

ResearchDGX agent

arXiv:2606.00716v1 Announce Type: new Abstract: Inference and control in engineered physical systems pay a heavy physics cost at deployment: state estimators, inverse-problem solvers, model-predictive

Hard Labels In! Rethinking the Role of Hard Labels in Mitigating Local Semantic Drift

SafetyDGX agent

arXiv:2512.15647v3 Announce Type: replace Abstract: Soft labels from teacher models are a de facto practice for knowledge transfer and large-scale dataset distillation (e.g., SRe2L, LPLD). However, wh

HASTE: Hardware-Aware Dynamic Sparse Training for Large Output Spaces

SafetyDGX agent

arXiv:2606.01117v1 Announce Type: cross Abstract: Extreme multi-label classification (XMC) involves learning models over large output spaces with millions of labels, making the output layer a memory-c

Holo3.1: Fast & Local Computer Use Agents

ToolsDGX agent

Holo3.1 is a computer use agent developed by Hugging Face that enables fast, local execution of tasks on computing systems without requiring cloud infrastructure. The model is designed to interpret an

Honey, I Shrunk the Arc de Triomphe!

ResearchDGX agent

arXiv:2606.02379v1 Announce Type: new Abstract: Metric scale monocular geometry estimation has seen significant progress through large-scale data aggregation, yet current foundation models suffer from

How Far Do Auto-Interpretation Labels Generalize: A Controlled Study Across Languages, Scripts, and Rewordings

ResearchDGX agent

arXiv:2606.00356v1 Announce Type: new Abstract: Sparse autoencoder (SAE) features are increasingly used to interpret language models, with auto-generated natural-language labels serving as the primary

Hybrid Probabilistic Forecasting of Under-Five Malaria Admissions in Ghana: A Gaussian Process Regression with Holt-Winters Smoothing

ResearchDGX agent

arXiv:2606.00834v1 Announce Type: cross Abstract: Accurate malaria forecasting remains a major challenge in sub-Saharan Africa, where strong seasonality, reporting uncertainty, and non-stationary tran

If your daughter needs tutoring in algebra, you can probably find someone cheaper than Albert Einstein. Giving every task to GPT5.5 or Opus …

IndustryDGX agent

If your daughter needs tutoring in algebra, you can probably find someone cheaper than Albert Einstein. Giving every task to GPT5.5 or Opus 4.8 is overkill. Often times you can get the task done just

Improving Visual Token Reduction via Rectifying Distortions for Efficient Multimodal LLM Inference

ResearchDGX agent

arXiv:2606.01711v1 Announce Type: new Abstract: Recent advancements in Multimodal Large Language Models (MLLMs) have achieved remarkable success in vision-language tasks, yet the quadratic computation

Information-Theoretic Lower Bounds for Bit-Constrained Stochastic Optimization via a Reduction to Compressed Gaussian Mean Estimation

ResearchDGX agent

arXiv:2606.00703v1 Announce Type: cross Abstract: Low-precision pretraining (FP8, MXFP4, NVFP4) is now standard for frontier language models, yet the literature is almost entirely achievability -- alg

← Previous
1…756757758759760…1018
Next →