AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,098 results
10 Apr 2026

FMI@SU ToxHabits: Evaluating LLMs Performance on Toxic Habit Extraction in Spanish Clinical Texts

Model ReleasesDGX agent

arXiv:2604.06403v1 Announce Type: cross Abstract: The paper presents an approach for the recognition of toxic habits named entities in Spanish clinical texts. The approach was developed for the ToxHab

FORGE:Fine-grained Multimodal Evaluation for Manufacturing Scenarios

Model ReleasesDGX agent

arXiv:2604.07413v1 Announce Type: new Abstract: The manufacturing sector is increasingly adopting Multimodal Large Language Models (MLLMs) to transition from simple perception to autonomous execution,

From Fragments to Facts: A Curriculum-Driven DPO Approach for Generating Hindi News Veracity Explanations

Model Releases

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2507.05179v4 Announce Type: replace Abstract: In an era of rampant misinformation, generating reliable news explanations is vital, especially for under-represented languages like Hindi. Lacking

GameWorld: Towards Standardized and Verifiable Evaluation of Multimodal Game Agents

Model ReleasesDGX agent

arXiv:2604.07429v1 Announce Type: new Abstract: Towards an embodied generalist for real-world interaction, Multimodal Large Language Model (MLLM) agents still suffer from challenging latency, sparse f

Gemma 4 on AI Gateway

Model ReleasesDGX agent

Google's Gemma 4 models — the 26B (Mixture of Experts) and 31B (Dense) variants — are now available on Vercel AI Gateway. Built on the same architecture as Gemini 3, both open models support functi...

Gemma 4:e4b offloads to RAM despite having just half of VRAM used.

Model ReleasesDGX agent

Users on the r/ollama subreddit reported that the **Gemma 4 E4B** model in Ollama offloads layers to RAM even when GPU VRAM is only partially utilized. This behavior is linked to how Ollama and lla...

Gen Z’s love-hate relationship with AI

Model ReleasesDGX agent

Gen Z is increasingly disillusioned with AI - just not enough to stop using it. A new Gallup report released this week, based on responses from nearly 1,600 people ages 14 to 29 across the US, suggest

Generating Attribution Reports for Manipulated Facial Images: A Dataset and Baseline

Model ReleasesDGX agent

arXiv:2412.19685v2 Announce Type: replace-cross Abstract: Existing facial forgery detection methods typically focus on binary classification or pixel-level localization, providing little semantic insi

GLM-5.1 by @Zai_org is now #3 in Code Arena - surpassing Gemini 3.1 and GPT-5.4, and now on par with Claude Sonnet 4.6. The first frontier l…

Model ReleasesDGX agent

GLM-5.1 by @Zai_org is now #3 in Code Arena - surpassing Gemini 3.1 and GPT-5.4, and now on par with Claude Sonnet 4.6. The first frontier level open model to break into the top 3. It’s a major +90 po

Graph Neural Networks for Misinformation Detection: Performance-Efficiency Trade-offs

Model ReleasesDGX agent

arXiv:2604.08131v1 Announce Type: new Abstract: The rapid spread of online misinformation has led to increasingly complex detection models, including large language models and hybrid architectures. Ho

GRASS: Gradient-based Adaptive Layer-wise Importance Sampling for Memory-efficient Large Language Model Fine-tuning

Model ReleasesDGX agent

arXiv:2604.07808v1 Announce Type: new Abstract: Full-parameter fine-tuning of large language models is constrained by substantial GPU memory requirements. Low-rank adaptation methods mitigate this cha

Grok 4.20 hitting 83% on non-hallucination. Values truth. Claude ~74%. Others sitting in the 60s… or way lower. Less guessing. More honesty …

Model ReleasesDGX agent

Grok 4.20 hitting 83% on non-hallucination. Values truth. Claude ~74%. Others sitting in the 60s… or way lower. Less guessing. More honesty when it doesn’t know. That’s a different kind of intelligenc

GroupGPT: A Token-efficient and Privacy-preserving Agentic Framework for Multi-User Chat Assistant

Model ReleasesDGX agent

arXiv:2603.01059v3 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have enabled increasingly capable chatbots. However, most existing systems focus on single-user sett

GS-Surrogate: Deformable Gaussian Splatting for Parameter Space Exploration of Ensemble Simulations

Model ReleasesDGX agent

arXiv:2604.06358v1 Announce Type: cross Abstract: Exploring ensemble simulations is increasingly important across many scientific domains. However, supporting flexible post-hoc exploration remains cha

haha cool, this worked. an MCP server that allows Claude to build it's own reusable skills https://github.com/yoheinakajima/selfMCP (1473 Lo…

Model ReleasesDGX agent

haha cool, this worked. an MCP server that allows Claude to build it's own reusable skills https://github.com/yoheinakajima/selfMCP (1473 LoC) basically a server with skills to CRUD skills in this vid

Harf-Speech: A Clinically Aligned Framework for Arabic Phoneme-Level Speech Assessment

Model ReleasesDGX agent

arXiv:2604.06191v1 Announce Type: cross Abstract: Automated phoneme-level pronunciation assessment is vital for scalable speech therapy and language learning, yet validated tools for Arabic remain sca

Help us grow our list of community middlewares! https://docs.langchain.com/oss/python/integrations/middleware#community-integrations

Model ReleasesDGX agent

Help us grow our list of community middlewares! https://docs.langchain.com/oss/python/integrations/middleware#community-integrations @IeloEmanuele is on a roll! you can now use claude code's advisor s

HiCI: Hierarchical Construction-Integration for Long-Context Attention

Model ReleasesDGX agent

arXiv:2603.20843v2 Announce Type: replace Abstract: Long-context language modeling is commonly framed as a scalability challenge of token-level attention, yet local-to-global information structuring r

HingeMem: Boundary Guided Long-Term Memory with Query Adaptive Retrieval for Scalable Dialogues

Model ReleasesDGX agent

arXiv:2604.06845v1 Announce Type: cross Abstract: Long-term memory is critical for dialogue systems that support continuous, sustainable, and personalized interactions. However, existing methods rely

HistDiT: A Structure-Aware Latent Conditional Diffusion Model for High-Fidelity Virtual Staining in Histopathology

Model ReleasesDGX agent

arXiv:2604.08305v1 Announce Type: cross Abstract: Immunohistochemistry (IHC) is essential for assessing specific immune biomarkers like Human Epidermal growth-factor Receptor 2 (HER2) in breast cancer

Holistic Optimal Label Selection for Robust Prompt Learning under Partial Labels

Model ReleasesDGX agent

arXiv:2604.06614v1 Announce Type: cross Abstract: Prompt learning has gained significant attention as a parameter-efficient approach for adapting large pre-trained vision-language models to downstream

How SAP Concur automates expense reporting with agentic AI

Model ReleasesDGX agent

For decades, expense automation relied on a simple premise: If the machine can read the text, it can do the work. But anyone who has ever tried to scan a crumpled, smudged, or sun-bleached receipt fro

https://docs.pinecone.io/integrations/gemini-cli

Model ReleasesDGX agent

I was unable to retrieve the specific content from the Pinecone Gemini CLI documentation page (`https://docs.pinecone.io/integrations/gemini-cli`) or the linked X (Twitter) post, as neither was ret...

HY-Embodied-0.5: Embodied Foundation Models for Real-World Agents

Model ReleasesDGX agent

arXiv:2604.07430v1 Announce Type: new Abstract: We introduce HY-Embodied-0.5, a family of foundation models specifically designed for real-world embodied agents. To bridge the gap between general Visi

HyperMem: Hypergraph Memory for Long-Term Conversations

Model ReleasesDGX agent

arXiv:2604.08256v1 Announce Type: new Abstract: Long-term memory is essential for conversational agents to maintain coherence, track persistent tasks, and provide personalized interactions across exte

I strongly suspect that Claude Mythos is a looped language model, as described in the paper 'Scaling Latent Reasoning via Looped Language Mo…

Model ReleasesDGX agent

I strongly suspect that Claude Mythos is a looped language model, as described in the paper 'Scaling Latent Reasoning via Looped Language Models' from ByteDance The authors of that paper called out gr

IatroBench: Pre-Registered Evidence of Iatrogenic Harm from AI Safety Measures

Model ReleasesDGX agent

arXiv:2604.07709v1 Announce Type: cross Abstract: Ask a frontier model how to taper six milligrams of alprazolam (psychiatrist retired, ten days of pills left, abrupt cessation causes seizures) and it

@IeloEmanuele is on a roll! you can now use claude code's advisor strategy with @LangChain agents because langchain/deepagents are provider …

Model ReleasesDGX agent

@IeloEmanuele is on a roll! you can now use claude code's advisor strategy with @LangChain agents because langchain/deepagents are provider agnostic, you can use different providers for your advisor a

if it creates a skill, and it errors, it will (at least sometimes) just try to fix it. what’s cool is that all the web search and code writi…

Model ReleasesDGX agent

if it creates a skill, and it errors, it will (at least sometimes) just try to fix it. what’s cool is that all the web search and code writing is offloaded to Claude (or whatever chat you’re using), a

If you want to see the setup... https://x.com/alliekmiller/status/2042728780847047131?s=20

Model ReleasesDGX agent

If you want to see the setup... https://x.com/alliekmiller/status/2042728780847047131?s=20 So many people wanted to see my knowledge management system in Claude code, so here it is. All you need is Ob

Illocutionary Explanation Planning for Source-Faithful Explanations in Retrieval-Augmented Language Models

Model ReleasesDGX agent

arXiv:2604.06211v1 Announce Type: cross Abstract: Natural language explanations produced by large language models (LLMs) are often persuasive, but not necessarily scrutable: users cannot easily verify

In a world where writing code to build websites and apps is trivial (thank you Lovable, Cursor, Claude,...), the real differentiation for yo…

Model ReleasesDGX agent

In a world where writing code to build websites and apps is trivial (thank you Lovable, Cursor, Claude,...), the real differentiation for you and your company (and what makes you successful) will be h

In-Context Decision Making for Optimizing Complex AutoML Pipelines

Model ReleasesDGX agent

arXiv:2508.13657v2 Announce Type: replace-cross Abstract: Combined Algorithm Selection and Hyperparameter Optimization (CASH) has been fundamental to traditional AutoML systems. However, with the adva

Information as Structural Alignment: A Dynamical Theory of Continual Learning

Model ReleasesDGX agent

arXiv:2604.07108v1 Announce Type: cross Abstract: Catastrophic forgetting is not an engineering failure. It is a mathematical consequence of storing knowledge as global parameter superposition. Existi

Infusion: Shaping Model Behavior by Editing Training Data via Influence Functions

Model ReleasesDGX agent

arXiv:2602.09987v5 Announce Type: replace-cross Abstract: Influence functions are commonly used to attribute model behavior to training documents. We explore the reverse: crafting training data that i

Initialisation Determines the Basin: Efficient Codebook Optimisation for Extreme LLM Quantization

Model ReleasesDGX agent

arXiv:2604.08118v1 Announce Type: new Abstract: Additive quantization enables extreme LLM compression with O(1) lookup-table dequantization, making it attractive for edge deployment. Yet at 2-bit prec

Instance-Adaptive Parametrization for Amortized Variational Inference

Model ReleasesDGX agent

arXiv:2604.06796v1 Announce Type: cross Abstract: Latent variable models, including variational autoencoders (VAE), remain a central tool in modern deep generative modeling due to their scalability an

InstAP: Instance-Aware Vision-Language Pre-Train for Spatial-Temporal Understanding

Model ReleasesDGX agent

arXiv:2604.08337v1 Announce Type: new Abstract: Current vision-language pre-training (VLP) paradigms excel at global scene understanding but struggle with instance-level reasoning due to global-only s

Invisible Influences: Investigating Implicit Intersectional Biases through Persona Engineering in Large Language Models

Model ReleasesDGX agent

arXiv:2604.06213v1 Announce Type: cross Abstract: Large Language Models (LLMs) excel at human-like language generation but often embed and amplify implicit, intersectional biases, especially under per

JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency

Model ReleasesDGX agent

arXiv:2604.03044v2 Announce Type: replace-cross Abstract: We introduce JoyAI-LLM Flash, an efficient Mixture-of-Experts (MoE) language model designed to redefine the trade-off between strong performan

k-Maximum Inner Product Attention for Graph Transformers and the Expressive Power of GraphGPS

Model ReleasesDGX agent

arXiv:2604.03815v2 Announce Type: replace-cross Abstract: Graph transformers have shown promise in overcoming limitations of traditional graph neural networks, such as oversquashing and difficulties i

k-server-bench: Automating Potential Discovery for the k-Server Conjecture

Model ReleasesDGX agent

arXiv:2604.07240v1 Announce Type: cross Abstract: We introduce a code-based challenge for automated, open-ended mathematical discovery based on the k-server conjecture, a central open problem in com

Kathleen: Oscillator-Based Byte-Level Text Classification Without Tokenization or Attention

Model ReleasesDGX agent

arXiv:2604.07969v1 Announce Type: new Abstract: We present Kathleen, a text classification architecture that operates directly on raw UTF-8 bytes using frequency-domain processing -- requiring no toke

KEO: Knowledge Extraction on OMIn via Knowledge Graphs and RAG for Safety-Critical Aviation Maintenance

Model ReleasesDGX agent

arXiv:2510.05524v2 Announce Type: replace Abstract: We present Knowledge Extraction on OMIn (KEO), a domain-specific knowledge extraction and reasoning framework with large language models (LLMs) in s

KITE: Keyframe-Indexed Tokenized Evidence for VLM-Based Robot Failure Analysis

Model ReleasesDGX agent

arXiv:2604.07034v1 Announce Type: cross Abstract: We present KITE, a training-free, keyframe-anchored, layout-grounded front-end that converts long robot-execution videos into compact, interpretable t

Knowledge Graphs Generation from Cultural Heritage Texts: Combining LLMs and Ontological Engineering for Scholarly Debates

Model ReleasesDGX agent

arXiv:2511.10354v1 Announce Type: cross Abstract: Cultural Heritage texts contain rich knowledge that is difficult to query systematically due to the challenges of converting unstructured discourse in

Kuramoto Oscillatory Phase Encoding: Neuro-inspired Synchronization for Improved Learning Efficiency

Model ReleasesDGX agent

arXiv:2604.07904v1 Announce Type: cross Abstract: Spatiotemporal neural dynamics and oscillatory synchronization are widely implicated in biological information processing and have been hypothesized t

KV Cache Offloading for Context-Intensive Tasks

Model ReleasesDGX agent

arXiv:2604.08426v1 Announce Type: cross Abstract: With the growing demand for long-context LLMs across a wide range of applications, the key-value (KV) cache has become a critical bottleneck for both

La France a parmi les meilleurs mathématiciens et ingénieurs IA du monde. On le sait. On les embauche partout ailleurs. Et la première chose…

Model ReleasesDGX agent

La France a parmi les meilleurs mathématiciens et ingénieurs IA du monde. On le sait. On les embauche partout ailleurs. Et la première chose qu'on fait au moment où on pourrait enfin capitaliser dessu

Learning Debt and Cost-Sensitive Bayesian Retraining: A Forecasting Operations Framework

Model ReleasesDGX agent

arXiv:2604.06438v1 Announce Type: cross Abstract: Forecasters often choose retraining schedules by convention rather than by an explicit decision rule. This paper gives that decision a posterior-space

Learning the Stellar Structure Equations via Self-supervised Physics-Informed Neural Networks

Model ReleasesDGX agent

arXiv:2604.06255v1 Announce Type: cross Abstract: Stellar astrophysics relies critically on accurate descriptions of the physical conditions inside stars. Traditional solvers such as exttt{MESA} (Mo

LifeAlign: Lifelong Alignment for Large Language Models with Memory-Augmented Focalized Preference Optimization

Model ReleasesDGX agent

arXiv:2509.17183v3 Announce Type: replace-cross Abstract: Alignment plays a crucial role in Large Language Models (LLMs) in aligning with human preferences on a specific task/domain. Traditional align

LiloDriver: A Lifelong Learning Framework for Closed-loop Motion Planning in Long-tail Autonomous Driving Scenarios

Model ReleasesDGX agent

arXiv:2505.17209v2 Announce Type: replace Abstract: Recent advances in autonomous driving research towards motion planners that are robust, safe, and adaptive. However, existing rule-based and data-dr

LiteParse is the best document parsing library for coding agents. It's free, fast, integrates natively with the LLM's native visual understa…

Model ReleasesDGX agent

LiteParse is the best document parsing library for coding agents. It's free, fast, integrates natively with the LLM's native visual understanding capabilities, and comes with support for 50+ formats a

llama.cpp now supports various small OCR models that can run on low-end devices. These models are small enough to run on GPU with 4GB VRAM, …

Model ReleasesDGX agent

llama.cpp now supports various small OCR models that can run on low-end devices. These models are small enough to run on GPU with 4GB VRAM, and some of them can even run on CPU with decent performance

LLM Spirals of Delusion: A Benchmarking Audit Study of AI Chatbot Interfaces

Model ReleasesDGX agent

arXiv:2604.06188v1 Announce Type: cross Abstract: People increasingly hold sustained, open-ended conversations with large language models (LLMs). Public reports and early studies suggest that, in such

LNN-PINN: A Unified Physics-Only Training Framework with Liquid Residual Blocks

Model ReleasesDGX agent

arXiv:2508.08935v4 Announce Type: replace Abstract: Physics-informed neural networks (PINNs) have attracted considerable attention for their ability to integrate partial differential equation priors i

LoFT: Parameter-Efficient Fine-Tuning for Long-tailed Semi-Supervised Learning in Open-World Scenarios

Model ReleasesDGX agent

arXiv:2509.09926v5 Announce Type: replace Abstract: Long-tailed semi-supervised learning (LTSSL) presents a formidable challenge where models must overcome the scarcity of tail samples while mitigatin

Logics-Parsing-Omni Technical Report

Model ReleasesDGX agent

arXiv:2603.09677v3 Announce Type: replace Abstract: Addressing the challenges of fragmented task definitions and the heterogeneity of unstructured data in multimodal parsing, this paper proposes the O

LongWriter-Zero: Mastering Ultra-Long Text Generation via Reinforcement Learning

Model ReleasesDGX agent

arXiv:2506.18841v3 Announce Type: replace-cross Abstract: Ultra-long generation by large language models (LLMs) is a widely demanded scenario, yet it remains a significant challenge due to their maxim

← Previous
1…361362363364365…369
Next →