AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,106 results
10 Apr 2026

Help us grow our list of community middlewares! https://docs.langchain.com/oss/python/integrations/middleware#community-integrations

Model ReleasesDGX agent

Help us grow our list of community middlewares! https://docs.langchain.com/oss/python/integrations/middleware#community-integrations @IeloEmanuele is on a roll! you can now use claude code's advisor s

HiCI: Hierarchical Construction-Integration for Long-Context Attention

Model ReleasesDGX agent

arXiv:2603.20843v2 Announce Type: replace Abstract: Long-context language modeling is commonly framed as a scalability challenge of token-level attention, yet local-to-global information structuring r

HingeMem: Boundary Guided Long-Term Memory with Query Adaptive Retrieval for Scalable Dialogues


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

arXiv:2604.06845v1 Announce Type: cross Abstract: Long-term memory is critical for dialogue systems that support continuous, sustainable, and personalized interactions. However, existing methods rely

HistDiT: A Structure-Aware Latent Conditional Diffusion Model for High-Fidelity Virtual Staining in Histopathology

Model ReleasesDGX agent

arXiv:2604.08305v1 Announce Type: cross Abstract: Immunohistochemistry (IHC) is essential for assessing specific immune biomarkers like Human Epidermal growth-factor Receptor 2 (HER2) in breast cancer

Holistic Optimal Label Selection for Robust Prompt Learning under Partial Labels

Model ReleasesDGX agent

arXiv:2604.06614v1 Announce Type: cross Abstract: Prompt learning has gained significant attention as a parameter-efficient approach for adapting large pre-trained vision-language models to downstream

How SAP Concur automates expense reporting with agentic AI

Model ReleasesDGX agent

For decades, expense automation relied on a simple premise: If the machine can read the text, it can do the work. But anyone who has ever tried to scan a crumpled, smudged, or sun-bleached receipt fro

HY-Embodied-0.5: Embodied Foundation Models for Real-World Agents

Model ReleasesDGX agent

arXiv:2604.07430v1 Announce Type: new Abstract: We introduce HY-Embodied-0.5, a family of foundation models specifically designed for real-world embodied agents. To bridge the gap between general Visi

HyperMem: Hypergraph Memory for Long-Term Conversations

Model ReleasesDGX agent

arXiv:2604.08256v1 Announce Type: new Abstract: Long-term memory is essential for conversational agents to maintain coherence, track persistent tasks, and provide personalized interactions across exte

I strongly suspect that Claude Mythos is a looped language model, as described in the paper 'Scaling Latent Reasoning via Looped Language Mo…

Model ReleasesDGX agent

I strongly suspect that Claude Mythos is a looped language model, as described in the paper 'Scaling Latent Reasoning via Looped Language Models' from ByteDance The authors of that paper called out gr

IatroBench: Pre-Registered Evidence of Iatrogenic Harm from AI Safety Measures

Model ReleasesDGX agent

arXiv:2604.07709v1 Announce Type: cross Abstract: Ask a frontier model how to taper six milligrams of alprazolam (psychiatrist retired, ten days of pills left, abrupt cessation causes seizures) and it

@IeloEmanuele is on a roll! you can now use claude code's advisor strategy with @LangChain agents because langchain/deepagents are provider …

Model ReleasesDGX agent

@IeloEmanuele is on a roll! you can now use claude code's advisor strategy with @LangChain agents because langchain/deepagents are provider agnostic, you can use different providers for your advisor a

if it creates a skill, and it errors, it will (at least sometimes) just try to fix it. what’s cool is that all the web search and code writi…

Model ReleasesDGX agent

if it creates a skill, and it errors, it will (at least sometimes) just try to fix it. what’s cool is that all the web search and code writing is offloaded to Claude (or whatever chat you’re using), a

If you want to see the setup... https://x.com/alliekmiller/status/2042728780847047131?s=20

Model ReleasesDGX agent

If you want to see the setup... https://x.com/alliekmiller/status/2042728780847047131?s=20 So many people wanted to see my knowledge management system in Claude code, so here it is. All you need is Ob

Illocutionary Explanation Planning for Source-Faithful Explanations in Retrieval-Augmented Language Models

Model ReleasesDGX agent

arXiv:2604.06211v1 Announce Type: cross Abstract: Natural language explanations produced by large language models (LLMs) are often persuasive, but not necessarily scrutable: users cannot easily verify

In a world where writing code to build websites and apps is trivial (thank you Lovable, Cursor, Claude,...), the real differentiation for yo…

Model ReleasesDGX agent

In a world where writing code to build websites and apps is trivial (thank you Lovable, Cursor, Claude,...), the real differentiation for you and your company (and what makes you successful) will be h

In-Context Decision Making for Optimizing Complex AutoML Pipelines

Model ReleasesDGX agent

arXiv:2508.13657v2 Announce Type: replace-cross Abstract: Combined Algorithm Selection and Hyperparameter Optimization (CASH) has been fundamental to traditional AutoML systems. However, with the adva

Information as Structural Alignment: A Dynamical Theory of Continual Learning

Model ReleasesDGX agent

arXiv:2604.07108v1 Announce Type: cross Abstract: Catastrophic forgetting is not an engineering failure. It is a mathematical consequence of storing knowledge as global parameter superposition. Existi

Infusion: Shaping Model Behavior by Editing Training Data via Influence Functions

Model ReleasesDGX agent

arXiv:2602.09987v5 Announce Type: replace-cross Abstract: Influence functions are commonly used to attribute model behavior to training documents. We explore the reverse: crafting training data that i

Initialisation Determines the Basin: Efficient Codebook Optimisation for Extreme LLM Quantization

Model ReleasesDGX agent

arXiv:2604.08118v1 Announce Type: new Abstract: Additive quantization enables extreme LLM compression with O(1) lookup-table dequantization, making it attractive for edge deployment. Yet at 2-bit prec

Instance-Adaptive Parametrization for Amortized Variational Inference

Model ReleasesDGX agent

arXiv:2604.06796v1 Announce Type: cross Abstract: Latent variable models, including variational autoencoders (VAE), remain a central tool in modern deep generative modeling due to their scalability an

InstAP: Instance-Aware Vision-Language Pre-Train for Spatial-Temporal Understanding

Model ReleasesDGX agent

arXiv:2604.08337v1 Announce Type: new Abstract: Current vision-language pre-training (VLP) paradigms excel at global scene understanding but struggle with instance-level reasoning due to global-only s

Invisible Influences: Investigating Implicit Intersectional Biases through Persona Engineering in Large Language Models

Model ReleasesDGX agent

arXiv:2604.06213v1 Announce Type: cross Abstract: Large Language Models (LLMs) excel at human-like language generation but often embed and amplify implicit, intersectional biases, especially under per

JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency

Model ReleasesDGX agent

arXiv:2604.03044v2 Announce Type: replace-cross Abstract: We introduce JoyAI-LLM Flash, an efficient Mixture-of-Experts (MoE) language model designed to redefine the trade-off between strong performan

k-Maximum Inner Product Attention for Graph Transformers and the Expressive Power of GraphGPS

Model ReleasesDGX agent

arXiv:2604.03815v2 Announce Type: replace-cross Abstract: Graph transformers have shown promise in overcoming limitations of traditional graph neural networks, such as oversquashing and difficulties i

k-server-bench: Automating Potential Discovery for the k-Server Conjecture

Model ReleasesDGX agent

arXiv:2604.07240v1 Announce Type: cross Abstract: We introduce a code-based challenge for automated, open-ended mathematical discovery based on the k-server conjecture, a central open problem in com

Kathleen: Oscillator-Based Byte-Level Text Classification Without Tokenization or Attention

Model ReleasesDGX agent

arXiv:2604.07969v1 Announce Type: new Abstract: We present Kathleen, a text classification architecture that operates directly on raw UTF-8 bytes using frequency-domain processing -- requiring no toke

KEO: Knowledge Extraction on OMIn via Knowledge Graphs and RAG for Safety-Critical Aviation Maintenance

Model ReleasesDGX agent

arXiv:2510.05524v2 Announce Type: replace Abstract: We present Knowledge Extraction on OMIn (KEO), a domain-specific knowledge extraction and reasoning framework with large language models (LLMs) in s

KITE: Keyframe-Indexed Tokenized Evidence for VLM-Based Robot Failure Analysis

Model ReleasesDGX agent

arXiv:2604.07034v1 Announce Type: cross Abstract: We present KITE, a training-free, keyframe-anchored, layout-grounded front-end that converts long robot-execution videos into compact, interpretable t

Knowledge Graphs Generation from Cultural Heritage Texts: Combining LLMs and Ontological Engineering for Scholarly Debates

Model ReleasesDGX agent

arXiv:2511.10354v1 Announce Type: cross Abstract: Cultural Heritage texts contain rich knowledge that is difficult to query systematically due to the challenges of converting unstructured discourse in

Kuramoto Oscillatory Phase Encoding: Neuro-inspired Synchronization for Improved Learning Efficiency

Model ReleasesDGX agent

arXiv:2604.07904v1 Announce Type: cross Abstract: Spatiotemporal neural dynamics and oscillatory synchronization are widely implicated in biological information processing and have been hypothesized t

KV Cache Offloading for Context-Intensive Tasks

Model ReleasesDGX agent

arXiv:2604.08426v1 Announce Type: cross Abstract: With the growing demand for long-context LLMs across a wide range of applications, the key-value (KV) cache has become a critical bottleneck for both

La France a parmi les meilleurs mathématiciens et ingénieurs IA du monde. On le sait. On les embauche partout ailleurs. Et la première chose…

Model ReleasesDGX agent

La France a parmi les meilleurs mathématiciens et ingénieurs IA du monde. On le sait. On les embauche partout ailleurs. Et la première chose qu'on fait au moment où on pourrait enfin capitaliser dessu

Learning Debt and Cost-Sensitive Bayesian Retraining: A Forecasting Operations Framework

Model ReleasesDGX agent

arXiv:2604.06438v1 Announce Type: cross Abstract: Forecasters often choose retraining schedules by convention rather than by an explicit decision rule. This paper gives that decision a posterior-space

Learning the Stellar Structure Equations via Self-supervised Physics-Informed Neural Networks

Model ReleasesDGX agent

arXiv:2604.06255v1 Announce Type: cross Abstract: Stellar astrophysics relies critically on accurate descriptions of the physical conditions inside stars. Traditional solvers such as exttt{MESA} (Mo

LifeAlign: Lifelong Alignment for Large Language Models with Memory-Augmented Focalized Preference Optimization

Model ReleasesDGX agent

arXiv:2509.17183v3 Announce Type: replace-cross Abstract: Alignment plays a crucial role in Large Language Models (LLMs) in aligning with human preferences on a specific task/domain. Traditional align

LiloDriver: A Lifelong Learning Framework for Closed-loop Motion Planning in Long-tail Autonomous Driving Scenarios

Model ReleasesDGX agent

arXiv:2505.17209v2 Announce Type: replace Abstract: Recent advances in autonomous driving research towards motion planners that are robust, safe, and adaptive. However, existing rule-based and data-dr

LiteParse is the best document parsing library for coding agents. It's free, fast, integrates natively with the LLM's native visual understa…

Model ReleasesDGX agent

LiteParse is the best document parsing library for coding agents. It's free, fast, integrates natively with the LLM's native visual understanding capabilities, and comes with support for 50+ formats a

llama.cpp now supports various small OCR models that can run on low-end devices. These models are small enough to run on GPU with 4GB VRAM, …

Model ReleasesDGX agent

llama.cpp now supports various small OCR models that can run on low-end devices. These models are small enough to run on GPU with 4GB VRAM, and some of them can even run on CPU with decent performance

LLM Spirals of Delusion: A Benchmarking Audit Study of AI Chatbot Interfaces

Model ReleasesDGX agent

arXiv:2604.06188v1 Announce Type: cross Abstract: People increasingly hold sustained, open-ended conversations with large language models (LLMs). Public reports and early studies suggest that, in such

LNN-PINN: A Unified Physics-Only Training Framework with Liquid Residual Blocks

Model ReleasesDGX agent

arXiv:2508.08935v4 Announce Type: replace Abstract: Physics-informed neural networks (PINNs) have attracted considerable attention for their ability to integrate partial differential equation priors i

LoFT: Parameter-Efficient Fine-Tuning for Long-tailed Semi-Supervised Learning in Open-World Scenarios

Model ReleasesDGX agent

arXiv:2509.09926v5 Announce Type: replace Abstract: Long-tailed semi-supervised learning (LTSSL) presents a formidable challenge where models must overcome the scarcity of tail samples while mitigatin

Logics-Parsing-Omni Technical Report

Model ReleasesDGX agent

arXiv:2603.09677v3 Announce Type: replace Abstract: Addressing the challenges of fragmented task definitions and the heterogeneity of unstructured data in multimodal parsing, this paper proposes the O

LongWriter-Zero: Mastering Ultra-Long Text Generation via Reinforcement Learning

Model ReleasesDGX agent

arXiv:2506.18841v3 Announce Type: replace-cross Abstract: Ultra-long generation by large language models (LLMs) is a widely demanded scenario, yet it remains a significant challenge due to their maxim

LoRA-DA: Data-Aware Initialization for Low-Rank Adaptation via Asymptotic Analysis

Model ReleasesDGX agent

arXiv:2510.24561v2 Announce Type: replace-cross Abstract: LoRA has become a widely adopted method for PEFT, and its initialization methods have attracted increasing attention. However, existing method

Lost in Cultural Translation: Do LLMs Struggle with Math Across Cultural Contexts?

Model ReleasesDGX agent

arXiv:2503.18018v2 Announce Type: replace Abstract: We demonstrate that large language models' (LLMs) mathematical reasoning is culturally sensitive: testing 14 models from Anthropic, OpenAI, Google,

LPM 1.0: Video-based Character Performance Model

Model ReleasesDGX agent

arXiv:2604.07823v1 Announce Type: new Abstract: Performance, the externalization of intent, emotion, and personality through visual, vocal, and temporal behavior, is what makes a character alive. Lear

Lumbermark: Resistant Clustering by Chopping Up Mutual Reachability Minimum Spanning Trees

Model ReleasesDGX agent

arXiv:2604.07143v1 Announce Type: new Abstract: We introduce Lumbermark, a robust divisive clustering algorithm capable of detecting clusters of varying sizes, densities, and shapes. Lumbermark iterat

Making MLLMs Blind: Adversarial Smuggling Attacks in MLLM Content Moderation

Model ReleasesDGX agent

arXiv:2604.06950v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) are increasingly being deployed as automated content moderators. Within this landscape, we uncover a critic

Making Room for AI: Multi-GPU Molecular Dynamics with Deep Potentials in GROMACS

Model ReleasesDGX agent

arXiv:2604.07276v1 Announce Type: cross Abstract: GROMACS is a de-facto standard for classical Molecular Dynamics (MD). The rise of AI-driven interatomic potentials that pursue near-quantum accuracy a

MARCH: Evaluating the Intersection of Ambiguity Interpretation and Multi-hop Inference

Model ReleasesDGX agent

arXiv:2509.22750v3 Announce Type: replace Abstract: Real-world multi-hop QA is naturally linked with ambiguity, where a single query can trigger multiple reasoning paths that require independent resol

Matrix Profile for Anomaly Detection on Multidimensional Time Series

Model ReleasesDGX agent

arXiv:2409.09298v2 Announce Type: replace-cross Abstract: The Matrix Profile (MP), a versatile tool for time series data mining, has been shown effective in time series anomaly detection (TSAD). This

Matrix Profile for Time-Series Anomaly Detection: A Reproducible Open-Source Benchmark on TSB-AD

Model ReleasesDGX agent

arXiv:2604.02445v2 Announce Type: replace Abstract: Matrix Profile (MP) methods are an interpretable and scalable family of distance-based methods for time-series anomaly detection, but strong benchma

MedConclusion: A Benchmark for Biomedical Conclusion Generation from Structured Abstracts

Model ReleasesDGX agent

arXiv:2604.06505v1 Announce Type: cross Abstract: Large language models (LLMs) are widely explored for reasoning-intensive research tasks, yet resources for testing whether they can infer scientific c

MedDialBench: Benchmarking LLM Diagnostic Robustness under Parametric Adversarial Patient Behaviors

Model ReleasesDGX agent

arXiv:2604.06846v1 Announce Type: cross Abstract: Interactive medical dialogue benchmarks have shown that LLM diagnostic accuracy degrades significantly when interacting with non-cooperative patients,

MF-GLaM: A multifidelity stochastic emulator using generalized lambda models

Model ReleasesDGX agent

arXiv:2507.10303v2 Announce Type: replace-cross Abstract: Stochastic simulators exhibit intrinsic stochasticity due to unobservable, uncontrollable, or unmodeled input variables, resulting in random o

MinerU2.5-Pro: Pushing the Limits of Data-Centric Document Parsing at Scale

Model ReleasesDGX agent

arXiv:2604.04771v2 Announce Type: replace-cross Abstract: Current document parsing methods advance primarily through model architecture innovation, while systematic engineering of training data remain

Mitigating Distribution Sharpening in Math RLVR via Distribution-Aligned Hint Synthesis and Backward Hint Annealing

Model ReleasesDGX agent

arXiv:2604.07747v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) can improve low-k reasoning accuracy while narrowing solution coverage on challenging math que

Mitigating Spurious Background Bias in Multimedia Recognition with Disentangled Concept Bottlenecks

Model ReleasesDGX agent

arXiv:2510.15770v3 Announce Type: replace Abstract: Concept Bottleneck Models (CBMs) enhance interpretability by predicting human-understandable concepts as intermediate representations. However, exis

Mitigating Visual Context Degradation in Large Multimodal Models: A Training-Free Decoupled Agentic Framework

Model ReleasesDGX agent

arXiv:2509.23322v2 Announce Type: replace Abstract: With the continuous expansion of Large Language Models (LLMs) and advances in reinforcement learning, LLMs have demonstrated exceptional reasoning c

Mixed-Initiative Context: Structuring and Managing Context for Human-AI Collaboration

Model ReleasesDGX agent

arXiv:2604.07121v1 Announce Type: cross Abstract: In the human-AI collaboration area, the context formed naturally through multi-turn interactions is typically flattened into a chronological sequence

← Previous
1…362363364365366…369
Next →