AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

85,202Total entries
1Added by human
85,201Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,047 results
4 Jun 2026

Call your Hermes Agent

AgentsDGX agent

This post likely announces or describes how to invoke or utilize a Hermes Agent, possibly related to function calling, tool use, or agentic capabilities within the Hermes model framework developed by

Cartridges at Scale: Training Modular KV Caches over Large Document Collections

HardwareDGX agent

arXiv:2606.04557v1 Announce Type: new Abstract: Large Language Models can reason over long contexts, yet prefilling millions of tokens is wasteful as much of the content remains static across queries.

CleanCodec: Efficient and Robust Speech Tokenization via Perceptually Guided Encoding

ResearchDGX agent

arXiv:2606.04418v1 Announce Type: cross Abstract: Neural audio codecs are a key component of speech processing pipelines, compressing audio into discrete tokens for downstream modeling. However, exist

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Constrained Adaptive Rejection Sampling

ResearchDGX agent

arXiv:2510.01902v2 Announce Type: replace Abstract: Language Models (LMs) are increasingly used in applications where generated outputs must satisfy strict semantic or syntactic constraints. Existing

Continuum Robot State Estimation with Actuation Uncertainty

ResearchDGX agent

arXiv:2601.04493v3 Announce Type: replace Abstract: Continuum robots are flexible, slender manipulators well suited for confined surgical environments. In these settings, unknown interaction forces an

CRAFT: Cost-aware Refinement And Front-aware Tuning of Prompts

ResearchDGX agent

arXiv:2606.04661v1 Announce Type: new Abstract: Prompts tuned for accuracy often grow long, raising inference cost on every model call. The best accuracy-cost trade-off depends on the task and the bud

DAR: Deontic Reasoning with Agentic Harnesses

AgentsDGX agent

arXiv:2606.05009v1 Announce Type: cross Abstract: Deontic reasoning is the task of answering questions by applying explicit rules and policies to case-specific facts, for example computing tax liabili

Decentralized EM Algorithm for Gaussian Mixtures under Data Heterogeneity and Partial Labeling

ResearchDGX agent

arXiv:2411.05591v2 Announce Type: replace-cross Abstract: We systematically study several network-based Expectation-Maximization (EM) algorithms for the Gaussian mixture model within decentralized fed

Drift-Augmented Scoring: Text-Derived Noise Robustness for Zero-Shot Audio-Language Classification

ResearchDGX agent

arXiv:2606.04844v1 Announce Type: cross Abstract: Contrastive audio-language models such as CLAP enable zero-shot audio classification: a sound is labelled by matching its embedding to text prompt emb

DSA: Dynamic Step Allocation for Fast Autoregressive Video Generation

HardwareDGX agent

arXiv:2606.04432v1 Announce Type: new Abstract: Video diffusion transformers have achieved state-of-the-art visual quality, but their high inference cost remains a major bottleneck for real-time appli

EpiFormer: Learning Antigen-Antibody Interactions for Epitope Prediction via Geometric Deep Learning

ResearchDGX agent

arXiv:2606.04154v1 Announce Type: cross Abstract: Antibodies neutralize foreign antigens by binding to specific surface regions called epitopes. Computational epitope prediction is critical for unders

EvalStop: Using World Feedback to Detect and Correct Reward Overoptimization in Multi-Tenant RLHF Platforms

ResearchDGX agent

arXiv:2606.04145v1 Announce Type: cross Abstract: Cloud LLM fine-tuning platforms increasingly serve RLHF workloads, where a learned reward model is optimized as a proxy for human quality. As Gao et a

Explaining a probabilistic prediction on the simplex with Shapley compositions

ResearchDGX agent

arXiv:2408.01382v3 Announce Type: replace Abstract: Originating in game theory, Shapley values are widely used for explaining a machine learning model's prediction by quantifying the contribution of e

Failed Reasoning Traces Tell You What Is Fixable (But Not by Reading Them)

ResearchDGX agent

arXiv:2606.05145v1 Announce Type: cross Abstract: When post-trained language models fail on reasoning problems, the common test-time-scaling response is to spend more compute on additional attempts, a

Fast & Faithful Function Vectors

ResearchDGX agent

arXiv:2606.05079v1 Announce Type: new Abstract: Function vectors (FVs) are task representations elicited during in-context learning that can be used to steer Large Language Models (LLMs). However, des

Finite-Iteration Local Dynamics and Warm Starts for Alternating Power Iteration in Spiked Tensor PCA

ResearchDGX agent

arXiv:2606.04065v1 Announce Type: cross Abstract: We study simultaneous alternating power iteration for fixed-order asymmetric rank-one spiked tensor models. Our main contribution is a finite-iteratio

Grok on Cloudflare

IndustryDGX agent

Grok on Cloudflare We're partnering with @xai to bring Grok to @Cloudflare AI Gateway. • Grok LLMs, audio, image, and video models are now available through AI Gateway • Billed directly through Cloudf

Identifying Gems from Roman RAPIDly

Local AiDGX agent

arXiv:2606.05103v1 Announce Type: cross Abstract: The Nancy Grace Roman Space Telescope (Roman), set for launch as early as September 2026, will conduct wide-field infrared imaging surveys with unprec

if you had 100k to invest in OpenAI and/or Anthropic IPOs which would you go for?

SafetyDGX agent

Gary Marcus discusses investment strategy between potential OpenAI and Anthropic IPOs, likely weighing factors such as the companies' technological capabilities, market positioning, business models, a

In-Context Graphical Inference

SafetyDGX agent

arXiv:2606.05042v1 Announce Type: cross Abstract: Marginal inference in discrete graphical models forces a choice between exactness and scalability: exact algorithms are intractable for high-treewidth

Intra-Modal Neighbors Never Lie: Rectifying Inter-Modal Noisy Correspondence via Graph-Based Intra-Modal Reasoning

ResearchDGX agent

arXiv:2606.04061v1 Announce Type: new Abstract: Large-scale web-harvested datasets have fueled the progress of cross-modal retrieval but inevitably suffer from noisy correspondence, which severely deg

LazyAttention: Efficient Retrieval-Augmented Generation with Deferred Positional Encoding

ResearchDGX agent

arXiv:2606.04302v1 Announce Type: new Abstract: Key-value (KV) caching accelerates inference of large language models (LLMs) by reusing past computations for generated tokens. Its importance becomes e

Mid-Think: Training-Free Intermediate-Budget Reasoning via Token-Level Triggers

ResearchDGX agent

arXiv:2601.07036v2 Announce Type: replace-cross Abstract: Hybrid reasoning language models are commonly controlled through high-level Think/No-think instructions to regulate reasoning behavior, yet we

Motion-Guided Causal Disentanglement for Robust Multi-View Cine Cardiac MRI Diagnosis

TutorialsDGX agent

arXiv:2606.04414v1 Announce Type: new Abstract: Multi-view cardiac magnetic resonance (CMR) imaging provides complementary anatomical information and is widely used for noninvasive disease assessment.

okay but need to save some $$$ for tokens

AgentsDGX agent

This post likely discusses cost optimization strategies for managing API token usage and expenses, particularly relevant to users of language models or LLM-based services. The author emphasizes the pr

Optimizing the Cost-Quality Tradeoff of Agentic Theorem Provers in Lean

AgentsDGX agent

arXiv:2606.04883v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used in workflows for generating formal proofs in Lean. These workflows often decompose problems into smal

ParetoPilot: Zero-Surrogate Offline Multi-Objective Optimization via Infer-Perturb-Guide Diffusion

TutorialsDGX agent

arXiv:2606.04468v1 Announce Type: cross Abstract: Offline multi-objective optimization (Offline MOO) aims to discover novel Pareto-optimal designs based on static datasets without expensive environmen

Potential-Guided Flow Matching for Vision-Language-Action Policy Improvement

SafetyDGX agent

arXiv:2606.04968v1 Announce Type: new Abstract: Large vision-language-action (VLA) policies are increasingly trained as conditional generative models over action chunks. Yet deployment produces mixed-

Rethinking Sales Lead Scoring with LLM-based Hierarchical Preference Ranking

SafetyDGX agent

arXiv:2606.04387v1 Announce Type: cross Abstract: Sales lead conversion in high-stakes domains (e.g., automotive, real estate) differs fundamentally from e-commerce recommendation due to prolonged dec

RL Excursions during Pre-Training: Re-examining Policy Optimization for LLM training

SafetyDGX agent

arXiv:2606.04272v1 Announce Type: new Abstract: The standard LLM training pipeline applies reinforcement learning (RL) only after pre-training and supervised fine-tuning (SFT). We question this status

Self-Distilled Policy Gradient

SafetyDGX agent

arXiv:2606.04036v1 Announce Type: new Abstract: On-policy self-distillation, where a language model conditions on privileged context to supervise its own generations, is a promising source of dense su

SemBlock: Semantic Boundary Dynamic Blocks for Diffusion LLMs

Local AiDGX agent

arXiv:2606.04964v1 Announce Type: new Abstract: Diffusion language models (DLMs) generate text through iterative denoising, and blockwise decoding improves their practicality by committing tokens in l

Sequential Data Poisoning in LLM Post-Training

ResearchDGX agent

arXiv:2606.04929v1 Announce Type: new Abstract: LLM post-training proceeds through multiple stages, e.g., supervised fine-tuning (SFT) followed by reinforcement learning from human feedback (RLHF) or

Testing Neural Networks via Bayesian-Guided Exploration of Decision Landscapes

SafetyDGX agent

arXiv:2606.04314v1 Announce Type: new Abstract: As neural networks are increasingly deployed in safety-critical domains, testing is essential to evaluate and improve their reliability. Existing testin

The Right Measure for Physics-Constrained Generation: A Co-Area Correction for Posterior-Consistent PDE Inverse Problems

SafetyDGX agent

arXiv:2606.04804v1 Announce Type: new Abstract: Generative models -- diffusion and flow matching -- are increasingly used to solve partial differential equation (PDE) inverse problems, enforcing the g

Uncertainty-Aware (Un)Supervised Few-Shot User Adaptation for On-Device Personalized Human Activity Recognition

Local AiDGX agent

arXiv:2606.04798v1 Announce Type: new Abstract: Sensor-based Human Activity Recognition (HAR) models often degrade on unseen users due to domain shifts caused by individual movement patterns and senso

Uncertainty Estimation using Variance-Gated Distributions

ResearchDGX agent

arXiv:2509.08846v2 Announce Type: replace-cross Abstract: Evaluation of per-sample uncertainty quantification from neural networks is essential for decision-making involving high-risk applications. A

UniFine: A Unified and Fine-grained Approach for Zero-shot Vision-Language Understanding

ResearchDGX agent

arXiv:2307.00862v3 Announce Type: replace-cross Abstract: Vision-language tasks, such as VQA, SNLI-VE, and VCR are challenging because they require the model's reasoning ability to understand the sema

Unifying Information-Theoretic and Pair-Counting Clustering Similarity

ResearchDGX agent

arXiv:2511.03000v2 Announce Type: replace-cross Abstract: Comparing clusterings is central to evaluating unsupervised models, yet the many existing similarity measures can produce widely divergent, so

v0.30.5

Local AiDGX agent

v0.30.5 is a release candidate version of Ollama, a local large language model runner. The v0.30 series features improved compatibility and performance using llama.cpp and augments the MLX engine on A

What Type of Inference is Active Inference?

SafetyDGX agent

arXiv:2606.04935v1 Announce Type: new Abstract: Active inference casts decision-making as inference, with the Expected Free Energy (EFE) unifying goal-directed and information-seeking behavior. Recent

You Only Train Once: Differentiable Subset Selection for Omics Data

ResearchDGX agent

arXiv:2512.17678v2 Announce Type: replace-cross Abstract: Selecting compact and informative gene subsets from single-cell transcriptomic data is essential for biomarker discovery, improving interpreta

3 Jun 2026

A Cartesian-3j Framework for Machine Learning Interatomic Potentials

SafetyDGX agent

arXiv:2512.16882v2 Announce Type: replace-cross Abstract: Machine learning interatomic potentials (MLIPs) have brought substantial gains in the extrapolation capability in computational chemistry. How

A Factorized Low-Rank RNN Framework for Uncovering Independent Neural Latent Dynamics and Connectivity

ResearchDGX agent

arXiv:2511.13899v2 Announce Type: replace-cross Abstract: Low-rank recurrent neural networks (lrRNNs) are a class of models that uncover low-dimensional latent dynamics underlying neural population ac

A Fast Screening Approach for High-dimensional Outcomes and High-dimensional Predictors

ResearchDGX agent

arXiv:2606.03018v1 Announce Type: cross Abstract: Modeling interactions among multimodal, high-dimensional data is intrinsically challenging due to ultra-high dimensionality and complex dependence str

A Measurement-Driven Digital Twin Architecture for Plant-Level Biomass Estimation and Growth Forecasting in Hydroponic Systems

ResearchDGX agent

arXiv:2606.02796v1 Announce Type: new Abstract: Alternatives to soil-based horticulture, such as hydroponics, have been developed to respond to food distribution concerns for dense urban centers. A ne

Act Like a Pathologist: Tissue-Aware Whole Slide Image Reasoning

ResearchDGX agent

arXiv:2603.00667v3 Announce Type: replace Abstract: Computational pathology has advanced rapidly in recent years, driven by domain-specific image encoders and growing interest in using vision-language

Adapting Noise to Data: Generative Flows from 1D Processes

ResearchDGX agent

arXiv:2510.12636v5 Announce Type: replace-cross Abstract: The default Gaussian latent in flow-based generative models poses challenges when learning certain distributions such as heavy-tailed ones. We

Adaptive Latent Agentic Reasoning

AgentsDGX agent

arXiv:2606.02871v1 Announce Type: cross Abstract: Large reasoning models improve performance by generating extended chain-of-thought (CoT) reasoning, but this behavior becomes inefficient when applied

Adding MCP Tools to Reachy Mini

AgentsDGX agent

This article describes how to integrate Model Context Protocol (MCP) tools with Reachy Mini, a small humanoid robot platform. It likely covers the technical steps for extending Reachy Mini's capabilit

Aletheia: What Makes RLVR For Code Verifiers Tick?

SafetyDGX agent

arXiv:2601.12186v3 Announce Type: replace-cross Abstract: Multi-domain thinking verifiers trained via Reinforcement Learning with Verifiable Rewards (RLVR) are a cornerstone of modern post-training. H

Aligning Data-Driven Predictors with Allocation: A Decision-Focused Approach to Survival Analysis

SafetyDGX agent

arXiv:2606.02671v1 Announce Type: cross Abstract: Machine learning predictors have become essential tools for guiding automated decision making. However, a major misalignment persists: predictive mode

AlphaEval: A Comprehensive and Efficient Evaluation Framework for Formula Alpha Mining

ResearchDGX agent

arXiv:2508.13174v2 Announce Type: replace Abstract: Formula alpha mining, which generates predictive signals from financial data, is critical for quantitative investment. Although various algorithmic

ASymPO: Asymmetric-Scale Policy Optimization for Asynchronous LLM Post-Training Without Behavior Information

SafetyDGX agent

arXiv:2606.03070v1 Announce Type: cross Abstract: Asynchronous reinforcement learning can improve language-model post-training throughput by decoupling response generation from policy optimization, bu

b9489

Local AiDGX agent

Release b9489 of llama.cpp includes updates to hidden_act mapping in llama-model.cpp, additions of granite embedding multilingual R2 models, and support for setting hidden_activation in GGUF files. Th

BAHSD: Bridging the Long-tail Gap via Adaptive Distillation in Black-box Sequential Recommendation

ResearchDGX agent

arXiv:2606.03091v1 Announce Type: cross Abstract: Sequential recommendation systems are widely adopted but often deployed as black-box APIs, which has driven recent interest in model extraction to rep

Built with Grok

IndustryDGX agent

'Built with Grok' is a post from Elon Musk on X platform promoting or highlighting something created using Grok, xAI's large language model. The post likely showcases an application, feature, or capab

Co-evolving Agent Architectures and Interpretable Reasoning for Automated Optimization

AgentsDGX agent

arXiv:2604.17708v2 Announce Type: replace Abstract: Automating operations research (OR) with large language models (LLMs) remains limited by hand-crafted reasoning--execution workflows. Complex OR tas

Coherence Maximization Improves Pluralistic Alignment

SafetyDGX agent

arXiv:2606.03110v1 Announce Type: new Abstract: Aligning AI systems with diverse human values requires value specifications grounded in concrete examples, but generating such examples without extensiv

CORE: Conflict-Oriented Reasoning for General Multimodal Manipulation Detection

ResearchDGX agent

arXiv:2606.03066v1 Announce Type: new Abstract: The rapid rise of generative AI has made multimodal fake news increasingly realistic and pervasive, posing severe threats to public trust and social sta

← Previous
1…754755756757758…1018
Next →