AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,106 results
Model Releases

If you want to see the setup... https://x.com/alliekmiller/status/2042728780847047131?s=20

DGX agent

If you want to see the setup... https://x.com/alliekmiller/status/2042728780847047131?s=20 So many people wanted to see my knowledge management system in Claude code, so here it is. All you need is Ob

model-releasesallie-k--miller--x
10 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Illocutionary Explanation Planning for Source-Faithful Explanations in Retrieval-Augmented Language Models

DGX agent

arXiv:2604.06211v1 Announce Type: cross Abstract: Natural language explanations produced by large language models (LLMs) are often persuasive, but not necessarily scrutable: users cannot easily verify

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

In a world where writing code to build websites and apps is trivial (thank you Lovable, Cursor, Claude,...), the real differentiation for yo…

DGX agent

In a world where writing code to build websites and apps is trivial (thank you Lovable, Cursor, Claude,...), the real differentiation for you and your company (and what makes you successful) will be h

model-releasesclem-delangue--x
10 Apr 2026
Model Releases

In-Context Decision Making for Optimizing Complex AutoML Pipelines

DGX agent

arXiv:2508.13657v2 Announce Type: replace-cross Abstract: Combined Algorithm Selection and Hyperparameter Optimization (CASH) has been fundamental to traditional AutoML systems. However, with the adva

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Information as Structural Alignment: A Dynamical Theory of Continual Learning

DGX agent

arXiv:2604.07108v1 Announce Type: cross Abstract: Catastrophic forgetting is not an engineering failure. It is a mathematical consequence of storing knowledge as global parameter superposition. Existi

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Infusion: Shaping Model Behavior by Editing Training Data via Influence Functions

DGX agent

arXiv:2602.09987v5 Announce Type: replace-cross Abstract: Influence functions are commonly used to attribute model behavior to training documents. We explore the reverse: crafting training data that i

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Initialisation Determines the Basin: Efficient Codebook Optimisation for Extreme LLM Quantization

DGX agent

arXiv:2604.08118v1 Announce Type: new Abstract: Additive quantization enables extreme LLM compression with O(1) lookup-table dequantization, making it attractive for edge deployment. Yet at 2-bit prec

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Instance-Adaptive Parametrization for Amortized Variational Inference

DGX agent

arXiv:2604.06796v1 Announce Type: cross Abstract: Latent variable models, including variational autoencoders (VAE), remain a central tool in modern deep generative modeling due to their scalability an

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

InstAP: Instance-Aware Vision-Language Pre-Train for Spatial-Temporal Understanding

DGX agent

arXiv:2604.08337v1 Announce Type: new Abstract: Current vision-language pre-training (VLP) paradigms excel at global scene understanding but struggle with instance-level reasoning due to global-only s

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Invisible Influences: Investigating Implicit Intersectional Biases through Persona Engineering in Large Language Models

DGX agent

arXiv:2604.06213v1 Announce Type: cross Abstract: Large Language Models (LLMs) excel at human-like language generation but often embed and amplify implicit, intersectional biases, especially under per

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency

DGX agent

arXiv:2604.03044v2 Announce Type: replace-cross Abstract: We introduce JoyAI-LLM Flash, an efficient Mixture-of-Experts (MoE) language model designed to redefine the trade-off between strong performan

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

k-Maximum Inner Product Attention for Graph Transformers and the Expressive Power of GraphGPS

DGX agent

arXiv:2604.03815v2 Announce Type: replace-cross Abstract: Graph transformers have shown promise in overcoming limitations of traditional graph neural networks, such as oversquashing and difficulties i

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

k-server-bench: Automating Potential Discovery for the k-Server Conjecture

DGX agent

arXiv:2604.07240v1 Announce Type: cross Abstract: We introduce a code-based challenge for automated, open-ended mathematical discovery based on the k-server conjecture, a central open problem in com

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Kathleen: Oscillator-Based Byte-Level Text Classification Without Tokenization or Attention

DGX agent

arXiv:2604.07969v1 Announce Type: new Abstract: We present Kathleen, a text classification architecture that operates directly on raw UTF-8 bytes using frequency-domain processing -- requiring no toke

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

KEO: Knowledge Extraction on OMIn via Knowledge Graphs and RAG for Safety-Critical Aviation Maintenance

DGX agent

arXiv:2510.05524v2 Announce Type: replace Abstract: We present Knowledge Extraction on OMIn (KEO), a domain-specific knowledge extraction and reasoning framework with large language models (LLMs) in s

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

KITE: Keyframe-Indexed Tokenized Evidence for VLM-Based Robot Failure Analysis

DGX agent

arXiv:2604.07034v1 Announce Type: cross Abstract: We present KITE, a training-free, keyframe-anchored, layout-grounded front-end that converts long robot-execution videos into compact, interpretable t

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Knowledge Graphs Generation from Cultural Heritage Texts: Combining LLMs and Ontological Engineering for Scholarly Debates

DGX agent

arXiv:2511.10354v1 Announce Type: cross Abstract: Cultural Heritage texts contain rich knowledge that is difficult to query systematically due to the challenges of converting unstructured discourse in

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Kuramoto Oscillatory Phase Encoding: Neuro-inspired Synchronization for Improved Learning Efficiency

DGX agent

arXiv:2604.07904v1 Announce Type: cross Abstract: Spatiotemporal neural dynamics and oscillatory synchronization are widely implicated in biological information processing and have been hypothesized t

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

KV Cache Offloading for Context-Intensive Tasks

DGX agent

arXiv:2604.08426v1 Announce Type: cross Abstract: With the growing demand for long-context LLMs across a wide range of applications, the key-value (KV) cache has become a critical bottleneck for both

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

La France a parmi les meilleurs mathématiciens et ingénieurs IA du monde. On le sait. On les embauche partout ailleurs. Et la première chose…

DGX agent

La France a parmi les meilleurs mathématiciens et ingénieurs IA du monde. On le sait. On les embauche partout ailleurs. Et la première chose qu'on fait au moment où on pourrait enfin capitaliser dessu

model-releasesyann-lecun--x
10 Apr 2026
Model Releases

Learning Debt and Cost-Sensitive Bayesian Retraining: A Forecasting Operations Framework

DGX agent

arXiv:2604.06438v1 Announce Type: cross Abstract: Forecasters often choose retraining schedules by convention rather than by an explicit decision rule. This paper gives that decision a posterior-space

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

Learning the Stellar Structure Equations via Self-supervised Physics-Informed Neural Networks

DGX agent

arXiv:2604.06255v1 Announce Type: cross Abstract: Stellar astrophysics relies critically on accurate descriptions of the physical conditions inside stars. Traditional solvers such as exttt{MESA} (Mo

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

LifeAlign: Lifelong Alignment for Large Language Models with Memory-Augmented Focalized Preference Optimization

DGX agent

arXiv:2509.17183v3 Announce Type: replace-cross Abstract: Alignment plays a crucial role in Large Language Models (LLMs) in aligning with human preferences on a specific task/domain. Traditional align

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

LiloDriver: A Lifelong Learning Framework for Closed-loop Motion Planning in Long-tail Autonomous Driving Scenarios

DGX agent

arXiv:2505.17209v2 Announce Type: replace Abstract: Recent advances in autonomous driving research towards motion planners that are robust, safe, and adaptive. However, existing rule-based and data-dr

model-releasesarxiv-cs-ro
10 Apr 2026
Model Releases

LiteParse is the best document parsing library for coding agents. It's free, fast, integrates natively with the LLM's native visual understa…

DGX agent

LiteParse is the best document parsing library for coding agents. It's free, fast, integrates natively with the LLM's native visual understanding capabilities, and comes with support for 50+ formats a

model-releasesjerry-liu--x
10 Apr 2026
Model Releases

llama.cpp now supports various small OCR models that can run on low-end devices. These models are small enough to run on GPU with 4GB VRAM, …

DGX agent

llama.cpp now supports various small OCR models that can run on low-end devices. These models are small enough to run on GPU with 4GB VRAM, and some of them can even run on CPU with decent performance

model-releasesgeorgi-gerganov--x
10 Apr 2026
Model Releases

LLM Spirals of Delusion: A Benchmarking Audit Study of AI Chatbot Interfaces

DGX agent

arXiv:2604.06188v1 Announce Type: cross Abstract: People increasingly hold sustained, open-ended conversations with large language models (LLMs). Public reports and early studies suggest that, in such

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

LNN-PINN: A Unified Physics-Only Training Framework with Liquid Residual Blocks

DGX agent

arXiv:2508.08935v4 Announce Type: replace Abstract: Physics-informed neural networks (PINNs) have attracted considerable attention for their ability to integrate partial differential equation priors i

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

LoFT: Parameter-Efficient Fine-Tuning for Long-tailed Semi-Supervised Learning in Open-World Scenarios

DGX agent

arXiv:2509.09926v5 Announce Type: replace Abstract: Long-tailed semi-supervised learning (LTSSL) presents a formidable challenge where models must overcome the scarcity of tail samples while mitigatin

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

Logics-Parsing-Omni Technical Report

DGX agent

arXiv:2603.09677v3 Announce Type: replace Abstract: Addressing the challenges of fragmented task definitions and the heterogeneity of unstructured data in multimodal parsing, this paper proposes the O

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

LongWriter-Zero: Mastering Ultra-Long Text Generation via Reinforcement Learning

DGX agent

arXiv:2506.18841v3 Announce Type: replace-cross Abstract: Ultra-long generation by large language models (LLMs) is a widely demanded scenario, yet it remains a significant challenge due to their maxim

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

LoRA-DA: Data-Aware Initialization for Low-Rank Adaptation via Asymptotic Analysis

DGX agent

arXiv:2510.24561v2 Announce Type: replace-cross Abstract: LoRA has become a widely adopted method for PEFT, and its initialization methods have attracted increasing attention. However, existing method

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Lost in Cultural Translation: Do LLMs Struggle with Math Across Cultural Contexts?

DGX agent

arXiv:2503.18018v2 Announce Type: replace Abstract: We demonstrate that large language models' (LLMs) mathematical reasoning is culturally sensitive: testing 14 models from Anthropic, OpenAI, Google,

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

LPM 1.0: Video-based Character Performance Model

DGX agent

arXiv:2604.07823v1 Announce Type: new Abstract: Performance, the externalization of intent, emotion, and personality through visual, vocal, and temporal behavior, is what makes a character alive. Lear

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Lumbermark: Resistant Clustering by Chopping Up Mutual Reachability Minimum Spanning Trees

DGX agent

arXiv:2604.07143v1 Announce Type: new Abstract: We introduce Lumbermark, a robust divisive clustering algorithm capable of detecting clusters of varying sizes, densities, and shapes. Lumbermark iterat

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

Making MLLMs Blind: Adversarial Smuggling Attacks in MLLM Content Moderation

DGX agent

arXiv:2604.06950v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) are increasingly being deployed as automated content moderators. Within this landscape, we uncover a critic

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Making Room for AI: Multi-GPU Molecular Dynamics with Deep Potentials in GROMACS

DGX agent

arXiv:2604.07276v1 Announce Type: cross Abstract: GROMACS is a de-facto standard for classical Molecular Dynamics (MD). The rise of AI-driven interatomic potentials that pursue near-quantum accuracy a

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

MARCH: Evaluating the Intersection of Ambiguity Interpretation and Multi-hop Inference

DGX agent

arXiv:2509.22750v3 Announce Type: replace Abstract: Real-world multi-hop QA is naturally linked with ambiguity, where a single query can trigger multiple reasoning paths that require independent resol

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Matrix Profile for Anomaly Detection on Multidimensional Time Series

DGX agent

arXiv:2409.09298v2 Announce Type: replace-cross Abstract: The Matrix Profile (MP), a versatile tool for time series data mining, has been shown effective in time series anomaly detection (TSAD). This

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Matrix Profile for Time-Series Anomaly Detection: A Reproducible Open-Source Benchmark on TSB-AD

DGX agent

arXiv:2604.02445v2 Announce Type: replace Abstract: Matrix Profile (MP) methods are an interpretable and scalable family of distance-based methods for time-series anomaly detection, but strong benchma

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

MedConclusion: A Benchmark for Biomedical Conclusion Generation from Structured Abstracts

DGX agent

arXiv:2604.06505v1 Announce Type: cross Abstract: Large language models (LLMs) are widely explored for reasoning-intensive research tasks, yet resources for testing whether they can infer scientific c

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

MedDialBench: Benchmarking LLM Diagnostic Robustness under Parametric Adversarial Patient Behaviors

DGX agent

arXiv:2604.06846v1 Announce Type: cross Abstract: Interactive medical dialogue benchmarks have shown that LLM diagnostic accuracy degrades significantly when interacting with non-cooperative patients,

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

MF-GLaM: A multifidelity stochastic emulator using generalized lambda models

DGX agent

arXiv:2507.10303v2 Announce Type: replace-cross Abstract: Stochastic simulators exhibit intrinsic stochasticity due to unobservable, uncontrollable, or unmodeled input variables, resulting in random o

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

MinerU2.5-Pro: Pushing the Limits of Data-Centric Document Parsing at Scale

DGX agent

arXiv:2604.04771v2 Announce Type: replace-cross Abstract: Current document parsing methods advance primarily through model architecture innovation, while systematic engineering of training data remain

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Mitigating Distribution Sharpening in Math RLVR via Distribution-Aligned Hint Synthesis and Backward Hint Annealing

DGX agent

arXiv:2604.07747v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) can improve low-k reasoning accuracy while narrowing solution coverage on challenging math que

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Mitigating Spurious Background Bias in Multimedia Recognition with Disentangled Concept Bottlenecks

DGX agent

arXiv:2510.15770v3 Announce Type: replace Abstract: Concept Bottleneck Models (CBMs) enhance interpretability by predicting human-understandable concepts as intermediate representations. However, exis

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Mitigating Visual Context Degradation in Large Multimodal Models: A Training-Free Decoupled Agentic Framework

DGX agent

arXiv:2509.23322v2 Announce Type: replace Abstract: With the continuous expansion of Large Language Models (LLMs) and advances in reinforcement learning, LLMs have demonstrated exceptional reasoning c

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Mixed-Initiative Context: Structuring and Managing Context for Human-AI Collaboration

DGX agent

arXiv:2604.07121v1 Announce Type: cross Abstract: In the human-AI collaboration area, the context formed naturally through multi-turn interactions is typically flattened into a chronological sequence

model-releasesarxiv-cs-ai
10 Apr 2026
← Previous
1…453454455456457…461
Next →