AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

85,202Total entries
1Added by human
85,201Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,805 results
Model Releases

iLoRA: Bayesian Low-Rank Adaptation with Latent Interaction Graphs for Microbiome Diagnosis

DGX agent

arXiv:2605.30179v1 Announce Type: cross Abstract: Parameter-efficient adaptation has made LLMs practical for domain prediction, but standard LoRA still relies on a static low-rank update and does not

model-releasesarxiv-cs-ai
29 May 2026
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

I’m going to let you in on a secret. I made a mistake once.* In August 2024.* It’s true. I actually got something wrong. I predicted that th…

DGX agent

I’m going to let you in on a secret. I made a mistake once.* In August 2024.* It’s true. I actually got something wrong. I predicted that there would be a collapse of the AI bubble (which in my judgem

model-releasesgary-marcus--x
29 May 2026
Model Releases

I'm suspicious of that that whole story about Uber blowing their AI budget and being disappointed in the results - I dug into it and it appe…

DGX agent

Simon Willison expresses skepticism about reports claiming Uber overspent on AI and was disappointed with the results, indicating he investigated the story and found issues with its accuracy or framin

model-releasessimon-willison--x
29 May 2026
Model Releases

Improving Adversarial Robustness of Attribution via Implicit Regularization

DGX agent

arXiv:2605.29983v1 Announce Type: cross Abstract: The adversarial robustness of attributions is a fundamental requirement for reliable explainability in deep learning, yet existing approaches typicall

model-releasesarxiv-cs-cv
29 May 2026
Model Releases

Improving Full Waveform Inversion in Large Model Era

DGX agent

arXiv:2603.00377v2 Announce Type: replace Abstract: Full Waveform Inversion (FWI) is a highly nonlinear and ill-posed problem that aims to recover subsurface velocity maps from surface-recorded seismi

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

In conversation with OpenAI’s @markchen90, Terence reflects on a future where AI reduces the cognitive friction of research, helps preserve …

DGX agent

In conversation with OpenAI’s @markchen90, Terence reflects on a future where AI reduces the cognitive friction of research, helps preserve the paths behind discovery, and expands what mathematicians

model-releasesopenai--x
29 May 2026
Model Releases

Indexing the Unreadable: LLM-Native Recursive Construction and Search of Service Taxonomies

DGX agent

arXiv:2605.29270v1 Announce Type: new Abstract: The era of the Internet of Agents (IoA) is taking shape: LLM agents are expected to fulfill user goals by orchestrating fast-growing populations of Mode

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Inferring the Size of Large Language Models From Popular Text Memorization

DGX agent

arXiv:2605.29223v1 Announce Type: new Abstract: The parameter counts of the most widely used large language models (LLMs) are often withheld by their developers, leaving model size -- a primary refere

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

Inform, Coach, Relate, Listen: Auditing LLM Caregiving Support Roles

DGX agent

arXiv:2605.29473v1 Announce Type: cross Abstract: Language models are increasingly being deployed for conversational support in informal caregiving contexts, where interactions often extend beyond inf

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

InsightEval: An Expert-Curated Benchmark for Assessing Insight Discovery in LLM-Driven Data Agents

DGX agent

arXiv:2511.22884v2 Announce Type: replace Abstract: Data analysis has become an indispensable part of scientific research. To discover the latent knowledge and insights hidden within massive datasets,

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Interesting that the GPT-5 Pro series models have consistently been the best models for single-shot attempts at the hardest problems since l…

DGX agent

Interesting that the GPT-5 Pro series models have consistently been the best models for single-shot attempts at the hardest problems since last summer. There has been no real competition in all that t

model-releasesethan-mollick--x
29 May 2026
Model Releases

Internal Representation, Not Clinical Knowledge: Where Apparent LLM Triage Failures Originate

DGX agent

arXiv:2605.29889v1 Announce Type: cross Abstract: Patient-voiced clinical-triage benchmarks report high under-triage rates for consumer LLMs for constrained multiple-choice output, yet the same cases

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Is Your LLM Overcharging You? Tokenization, Transparency, and Incentives

DGX agent

arXiv:2505.21627v4 Announce Type: replace-cross Abstract: State-of-the-art large language models require specialized hardware and substantial energy to operate. As a consequence, cloud-based services

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

I've largely switched over to using GPT-5.5 in recent weeks, which I like nearly as much as Opus 4.6 and 4.7, and is *very* reasonably price…

DGX agent

Jeremy Howard expresses positive views on GPT-5.5, stating he has recently switched to using it as his primary model and finds it nearly comparable to Anthropic's Opus 4.6 and 4.7 while offering signi

model-releasesjeremy-howard--x
29 May 2026
Model Releases

Just dropped 🧑‍🍳 Fixed version of DeepSeek-V4-Pro-NVFP4 by @NVIDIAAI https://huggingface.co/nvidia/DeepSeek-V4-Pro-NVFP4

DGX agent

NVIDIA has released a corrected version of DeepSeek-V4-Pro-NVFP4, a quantized model variant optimized for NVIDIA hardware using NV-FP4 (NVIDIA's 4-bit floating-point format). The model is available on

model-releasesclem-delangue--x
29 May 2026
Model Releases

K-FinHallu: A Hallucination Detection Benchmark for Multi-Turn RAG in Korean Finance

DGX agent

arXiv:2605.29523v1 Announce Type: new Abstract: Large Language Models (LLMs) have advanced financial automation through Retrieval-Augmented Generation (RAG), yet hallucinations remain a critical barri

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

KBF: Knowledge Boundary as Fingerprint for Language Model and Black-Box API Auditing

DGX agent

arXiv:2605.29524v1 Announce Type: cross Abstract: Relay and reseller APIs increasingly intermediate access to large language models (LLMs), but users have no direct way to verify that a claimed endpoi

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Kernel-based potential mean-field games with unbiased random Fourier U-statistics

DGX agent

arXiv:2605.29371v1 Announce Type: cross Abstract: We study the subclass of potential mean-field games in which the running interaction cost and the terminal target cost are both expressed through repr

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

Kernel Renormalization in Bayesian Deep Neural Networks: the Equivalent Wishart Ansatz in the Proportional Regime

DGX agent

arXiv:2605.29684v1 Announce Type: new Abstract: The scaling limit where both the size of the training set P and the width N of a deep neural network grow at the same rate, the so-called proportional-w

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

Knowledge Offloading: Decomposing LLMs into Sparse Backbones and Memory Modules

DGX agent

arXiv:2605.29075v1 Announce Type: new Abstract: LLMs encode both general capabilities and domain-specific knowledge in a single set of parameters. We ask whether this capacity can be reorganized: keep

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

Kronecker Embeddings: Byte-Level Structured Token Representations for Parameter-Efficient Language Models

DGX agent

arXiv:2605.29459v1 Announce Type: new Abstract: Large language models route every input through a learned embedding table of shape |V| x d_model, consuming hundreds of millions to billions of trainabl

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

Label-Free Reinforcement Learning via Cross-Model Entropy

DGX agent

arXiv:2605.29009v1 Announce Type: cross Abstract: Post-training large language models with reinforcement learning is bottlenecked by the reward signal. Existing approaches require either ground-truth

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Label Over Logic? How Source Cues Bias Human Fallacy Judgments More Than LLMs

DGX agent

arXiv:2605.29928v1 Announce Type: cross Abstract: As AI-generated and AI-assisted content floods online spaces, source labels attached to such content can distort human reasoning judgments, with downs

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Large language models reorganize representational geometry during in-context learning

DGX agent

arXiv:2605.28854v1 Announce Type: new Abstract: Large language models (LLMs) exhibit remarkable flexibility: they can adapt to novel tasks from in-context examples without any parameter updates, a cap

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

Latent Performance Profiling of Large Language Models

DGX agent

arXiv:2605.30018v1 Announce Type: new Abstract: Large language models (LLMs) frequently achieve impressive scores on standardized benchmarks, yet accuracy alone offers a limited view of their capabili

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

Leak@k: Unlearning Does Not Make LLMs Forget Under Probabilistic Decoding

DGX agent

arXiv:2511.04934v3 Announce Type: replace Abstract: Unlearning in large language models (LLMs) is critical for regulatory compliance and for building ethical generative AI systems that avoid producing

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

Learning Design Skills as Memory Policies for Agentic Photonic Inverse Design

DGX agent

arXiv:2605.29421v1 Announce Type: new Abstract: Photonic crystal fiber (PCF) inverse design remains challenging because candidate geometries must satisfy coupled optical targets under expensive electr

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

Learning Representations from 3D Gaussian Splats

DGX agent

arXiv:2605.29549v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) is a recent approach for scene rendering. Although primarily designed for view synthesis, its potential for scene understan

model-releasesarxiv-cs-cv
29 May 2026
Model Releases

Learning to Extrapolate to New Tasks: A Relational Approach to Task Extrapolation

DGX agent

arXiv:2605.30132v1 Announce Type: new Abstract: Modern learning systems excel at interpolation but struggle to generalize to unseen tasks outside the training distribution's support. This failure occu

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

Less is Enough: Synthesizing Diverse Data in LLM Feature Space with Sparse Autoencoders

DGX agent

arXiv:2602.10388v3 Announce Type: replace-cross Abstract: The diversity of post-training data is critical for effective downstream performance in large language models (LLMs). Many existing approaches

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Leveraging Routing Dynamics in Mixture-of-Experts Models for Efficient Language Adaptation

DGX agent

arXiv:2605.29714v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models are widely used to scale language models, yet their expert routing behavior and adaptation in a multilingual setting rem

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

libhmm: A Modern C++20 Library for Hidden Markov Models with Correct MLE Emission M-Steps

DGX agent

arXiv:2605.29208v1 Announce Type: cross Abstract: We describe libhmm, a C++20 library for Hidden Markov Model parameter estimation, sequence decoding, and model selection. libhmm addresses two gaps in

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

LiteCoder-Terminal: Scaling Long-Horizon Terminal Environments for Learning Language Agents

DGX agent

arXiv:2605.29559v1 Announce Type: new Abstract: Mastering terminal environments requires language agents capable of multi-step planning, feedback-grounded execution, and dynamic state adaptation. Howe

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

LiveSVG: Zero-Shot SVG Animation via Video Generation

DGX agent

arXiv:2605.30174v1 Announce Type: new Abstract: We introduce LiveSVG, a zero-shot approach for generating Scalable Vector Graphics (SVG) animations using video diffusion models. Current SVG animation

model-releasesarxiv-cs-cv
29 May 2026
Model Releases

llama.cpp now has an official website: https://llama.app Our goal is to make local AI accessible to everyone, and improving the user experie…

DGX agent

llama.cpp now has an official website: https://llama.app Our goal is to make local AI accessible to everyone, and improving the user experience is a big part of that. On the new landing page you’ll fi

model-releasesclem-delangue--x
29 May 2026
Model Releases

LLM-Evolved Domain-Independent Heuristics for Symbolic AI Planning

DGX agent

arXiv:2605.29649v1 Announce Type: new Abstract: Heuristic search is the dominant paradigm in symbolic AI planning, and the strongest heuristics are the result of decades of work by planning researcher

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

LoCoT2V-Bench: Benchmarking Long-Form and Complex Text-to-Video Generation

DGX agent

arXiv:2510.26412v3 Announce Type: replace-cross Abstract: Recent advances in text-to-video generation have achieved impressive performance on short clips, yet evaluating long-form generation under com

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

LogDx-CI: Benchmarking Log Reduction Tools for LLM Root-Cause Diagnosis

DGX agent

arXiv:2605.28876v1 Announce Type: cross Abstract: CI failure logs are large (median 5k lines, max 200k in this corpus) and noisy. Coding agents that try to debug them depend on an upstream tool to red

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Long-Context Modeling with Dynamic Hierarchical Sparse Attention for Memory-Constrained LLM Inference

DGX agent

arXiv:2510.24606v2 Announce Type: replace Abstract: The quadratic cost of attention limits the scalability of long-context LLMs, especially under limited hardware memory budgets. While attention is of

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

Loong: A Human-Like Long Document Translation Agent with Observe-and-Act Adaptive Context Selection

DGX agent

arXiv:2605.30274v1 Announce Type: cross Abstract: Document-level translation remains one of the most challenging tasks for large language models, which are constrained by limited context windows that

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

LoopFM: Learning frOm HistOrical RePresentations of Foundation Model for Recommendation

DGX agent

arXiv:2605.29280v1 Announce Type: cross Abstract: Knowledge distillation (KD) transfers a single scalar prediction from a large foundation model (FM) to compact vertical models (VMs), suffering from d

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

love it! @ggerganov 's llama.cpp delivering!

DGX agent

love it! @ggerganov 's llama.cpp delivering! pibot is now running fully local, using parakeet for STT, qwen3-tts for TTS, and Qwen 3.6 as the local multi-modal LLM via llama.cpp. The STT and TTS infer

model-releasesclem-delangue--x
29 May 2026
Model Releases

LUMINA: A Multi-Vendor Mammography Benchmark with Energy Harmonization Protocol

DGX agent

arXiv:2603.14644v3 Announce Type: replace-cross Abstract: Publicly available full-field digital mammography (FFDM) datasets remain limited in size, clinical annotations, and vendor diversity, hinderin

model-releasesarxiv-cs-cv
29 May 2026
Model Releases

MAGA-Bench: Machine-Augment-Generated Text via Alignment Detection Benchmark

DGX agent

arXiv:2601.04633v2 Announce Type: replace Abstract: Machine-Generated Text (MGT) is becoming increasingly difficult to distinguish from Human-Written Text (HWT). This trend has exacerbated malicious a

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

MarginGate: Sparse Margin-Triggered Verification for Batch-Invariant LLM Inference

DGX agent

arXiv:2605.30218v1 Announce Type: new Abstract: Temperature-zero BF16 LLM inference is often treated as reproducible, yet the same request can emit different tokens when decoded alone or inside a larg

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

MATNet: Multi-Level Fusion Transformer-Based Model for Day-Ahead PV Generation Forecasting

DGX agent

arXiv:2306.10356v3 Announce Type: replace-cross Abstract: Accurate forecasting of renewable generation is crucial to facilitate the integration of Renewable Energy Sources into the power system. Focus

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Mechanistic origins of catastrophic forgetting: why RL preserves circuits better than SFT?

DGX agent

arXiv:2605.28860v1 Announce Type: cross Abstract: Fine-tuning large language models (LLMs) frequently induces catastrophic forgetting of prior capabilities. Recent work has shown that reinforcement le

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

MechELK: A Mechanistic Interpretability Framework for Eliciting Latent Knowledge in Large Language Models

DGX agent

arXiv:2605.28825v1 Announce Type: new Abstract: Large language models (LLMs) frequently encode factual and reasoning knowledge in their internal representations that is not faithfully reflected in the

model-releasesarxiv-cs-cl
29 May 2026
← Previous
1…246247248249250…476
Next →