AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,037 results
15 Apr 2026

Safe-FedLLM: Delving into the Safety of Federated Large Language Models

Local AiDGX agent

arXiv:2601.07177v2 Announce Type: replace-cross Abstract: Federated learning (FL) addresses privacy and data-silo issues in the training of large language models (LLMs). Most prior work focuses on imp

SCRIPT: A Subcharacter Compositional Representation Injection Module for Korean Pre-Trained Language Models

ResearchDGX agent

arXiv:2604.12377v1 Announce Type: cross Abstract: Korean is a morphologically rich language with a featural writing system in which each character is systematically composed of subcharacter units know

14 Apr 2026

A Minimal Model of Representation Collapse: Frustration, Stop-Gradient, and Dynamics

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research
DGX agent

arXiv:2604.09979v1 Announce Type: cross Abstract: Self-supervised representation learning is central to modern machine learning because it extracts structured latent features from unlabeled data and e

Analytical Modeling and Correction of Distance Error in Homography-Based Ground-Plane Mapping

ResearchDGX agent

arXiv:2604.10805v1 Announce Type: new Abstract: Accurate distance estimation from monocular cameras is essential for intelligent monitoring systems. In many deployments, image coordinates are mapped t

Assessing Model-Agnostic XAI Methods against EU AI Act Explainability Requirements

SafetyDGX agent

arXiv:2604.09628v1 Announce Type: cross Abstract: Explainable AI (XAI) has evolved in response to expectations and regulations, such as the EU AI Act, which introduces regulatory requirements on AI-po

Beyond A Fixed Seal: Adaptive Stealing Watermark in Large Language Models

ApplicationsDGX agent

arXiv:2604.10893v1 Announce Type: cross Abstract: Watermarking provides a critical safeguard for large language model (LLM) services by facilitating the detection of LLM-generated text. Correspondingl

Characterizing Performance-Energy Trade-offs of Large Language Models in Multi-Request Workflows

HardwareDGX agent

arXiv:2604.09611v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used in applications forming multi-request workflows like document summarization, search-based copilots,

Decoupled Generative Modeling for Human-Object Interaction Synthesis

ResearchDGX agent

arXiv:2512.19049v2 Announce Type: replace Abstract: Synthesizing realistic human-object interaction (HOI) is essential for 3D computer vision and robotics, underpinning animation and embodied control.

Differentiable free energy surface: a variational approach to directly observing rare events using generative deep-learning models

SafetyDGX agent

arXiv:2604.09769v1 Announce Type: cross Abstract: Rare events are central to the evolution of complex many-body systems, characterized as key transitional configurations on the free energy surface (FE

FlowCoMotion: Text-to-Motion Generation via Token-Latent Flow Modeling

SafetyDGX agent

arXiv:2604.11083v1 Announce Type: cross Abstract: Text-to-motion generation is driven by learning motion representations for semantic alignment with language. Existing methods rely on either continuou

How some mathematicians are exploring ways to incorporate LLM models into their research without losing direct experience with mathematical understanding (Konstantin Kakaes/Quanta Magazine)

IndustryDGX agent

Konstantin Kakaes / Quanta Magazine: How some mathematicians are exploring ways to incorporate LLM models into their research without losing direct experience with mathematical understanding — Those c

Introducing Kernels on the Hugging Face Hub ✨ What if shipping a GPU kernel was as easy as pushing a model? - Pre-compiled for your exact GP…

HardwareDGX agent

Introducing Kernels on the Hugging Face Hub ✨ What if shipping a GPU kernel was as easy as pushing a model? - Pre-compiled for your exact GPU, PyTorch & OS - Multiple kernel versions coexist in one pr

Large Language Model as An Operator: An Experience-Driven Solution for Distribution Network Voltage Control

SafetyDGX agent

arXiv:2507.14800v2 Announce Type: replace-cross Abstract: With the advanced reasoning, contextual understanding, and information synthesis capabilities of large language models (LLMs), a novel paradig

LiveGesture Streamable Co-Speech Gesture Generation Model

ResearchDGX agent

arXiv:2604.10927v1 Announce Type: new Abstract: We propose LiveGesture, the first fully streamable, speech-driven full-body gesture generation framework that operates with zero look-ahead and supports

MimicLM: Zero-Shot Voice Imitation through Autoregressive Modeling of Pseudo-Parallel Speech Corpora

SafetyDGX agent

arXiv:2604.11552v1 Announce Type: cross Abstract: Voice imitation aims to transform source speech to match a reference speaker's timbre and speaking style while preserving linguistic content. A straig

Modeling, Analysis and Activation of Planar Viscoelastically-combined Rimless Wheels

ResearchDGX agent

arXiv:2604.11295v1 Announce Type: new Abstract: This paper proposes novel passive-dynamic walkers formed by two cross-shaped frames and eight viscoelastic elements. Since it is a combination of two fo

MSA: Memory Sparse Attention for Efficient End-to-End Memory Model Scaling to 100M Tokens

AgentsDGX agent

arXiv:2603.23516v2 Announce Type: replace-cross Abstract: Long-term memory is a cornerstone of human intelligence. Enabling AI to process lifetime-scale information remains a long-standing pursuit in

MSTN: A Lightweight and Fast Model for General TimeSeries Analysis

Local AiDGX agent

arXiv:2511.20577v3 Announce Type: replace Abstract: Real-world time series often exhibit strong non-stationarity, complex nonlinear dynamics, and behavior expressed across multiple temporal scales, fr

Rein3D: Reinforced 3D Indoor Scene Generation with Panoramic Video Diffusion Models

Local AiDGX agent

arXiv:2604.10578v1 Announce Type: new Abstract: The growing demand for Embodied AI and VR applications has highlighted the need for synthesizing high-quality 3D indoor scenes from sparse inputs. Howev

TInR: Exploring Tool-Internalized Reasoning in Large Language Models

SafetyDGX agent

arXiv:2604.10788v1 Announce Type: cross Abstract: Tool-Integrated Reasoning (TIR) has emerged as a promising direction by extending Large Language Models' (LLMs) capabilities with external tools durin

Towards Mitigating Modality Bias in Vision-Language Models for Temporal Action Localization

Local AiDGX agent

arXiv:2601.21078v3 Announce Type: replace Abstract: Temporal Action Localization (TAL) requires identifying both the boundaries and categories of actions in untrimmed videos. While vision-language mod

UBio-MolFM: A Universal Molecular Foundation Model for Bio-Systems

Local AiDGX agent

arXiv:2602.17709v2 Announce Type: replace-cross Abstract: All-atom molecular simulation serves as a quintessential ``computational microscope'' for understanding the machinery of life, yet it remains

UHD-GPGNet: UHD Video Denoising via Gaussian-Process-Guided Local Spatio-Temporal Modeling

Local AiDGX agent

arXiv:2604.11014v1 Announce Type: new Abstract: Ultra-high-definition (UHD) video denoising requires simultaneously suppressing complex spatio-temporal degradations, preserving fine textures and chrom

We achieved state-of-the-art performance in predicting which of 4.2 million genetic variants cause diseases by interpreting a genomics model…

ToolsDGX agent

We achieved state-of-the-art performance in predicting which of 4.2 million genetic variants cause diseases by interpreting a genomics model, in a new preprint with @MayoClinic. We're now releasing an

We are aware of issues impacting model availability via the Nous Portal and are working to restore service ASAP

ResearchDGX agent

Nous Research posted a status update on X acknowledging service disruptions affecting model availability through their Nous Portal. The announcement indicated the team was actively working to restore

13 Apr 2026

2026 AI Index Report: AI capability is accelerating, not plateauing, the US-China model gap has closed, the US leads in data centers and AI investment, and more (Stanford HAI)

IndustryDGX agent

Stanford HAI: 2026 AI Index Report: AI capability is accelerating, not plateauing, the US-China model gap has closed, the US leads in data centers and AI investment, and more — AI's influence on socie

Commanding Humanoid by Free-form Language: A Large Language Action Model with Unified Motion Vocabulary

SafetyDGX agent

arXiv:2511.22963v2 Announce Type: replace-cross Abstract: Enabling humanoid robots to follow free-form language commands is critical for seamless human-robot interaction, collaborative task execution,

Dual Mamba for Node-Specific Representation Learning: Tackling Over-Smoothing with Selective State Space Modeling

ResearchDGX agent

arXiv:2511.06756v3 Announce Type: replace Abstract: Over-smoothing remains a fundamental challenge in deep Graph Neural Networks (GNNs), where repeated message passing causes node representations to b

Fast Model-guided Instance-wise Adaptation Framework for Real-world Pansharpening with Fidelity Constraints

HardwareDGX agent

arXiv:2604.08903v1 Announce Type: new Abstract: Pansharpening aims to generate high-resolution multispectral (HRMS) images by fusing low-resolution multispectral (LRMS) and high-resolution panchromati

Investigating Multimodal Large Language Models to Support Usability Evaluation

ApplicationsDGX agent

arXiv:2508.16165v2 Announce Type: replace-cross Abstract: Usability evaluation is an essential method to support the design of effective and intuitive user interfaces (UIs). However, it commonly relie

Online Quantile Regression for Nonparametric Additive Models

ResearchDGX agent

arXiv:2604.08969v1 Announce Type: cross Abstract: This paper introduces a projected functional gradient descent algorithm (P-FGD) for training nonparametric additive quantile regression models in onli

SynDocDis: A Metadata-Driven Framework for Generating Synthetic Physician Discussions Using Large Language Models

ApplicationsDGX agent

arXiv:2604.08555v1 Announce Type: new Abstract: Physician-physician discussions of patient cases represent a rich source of clinical knowledge and reasoning that could feed AI agents to enrich and eve

We just OCR'd 27,000 arxiv papers into Markdown using an open 5B model, 16 parallel HF Jobs on L40S GPUs, and a mounted bucket. Total cost: …

IndustryDGX agent

We just OCR'd 27,000 arxiv papers into Markdown using an open 5B model, 16 parallel HF Jobs on L40S GPUs, and a mounted bucket. Total cost: $850 Total time: ~29 hours Jobs that crashed: 0 This now pow

12 Apr 2026

Let’s be honest… everyone was waiting to see how Tesla FSD 14.3 would handle Manhattan. @scotsrule08 and I took my 2026 HW4 Model Y out for …

IndustryDGX agent

Let’s be honest… everyone was waiting to see how Tesla FSD 14.3 would handle Manhattan. @scotsrule08 and I took my 2026 HW4 Model Y out for a 2-hour drive through NYC - and it seriously delivered. Thi

New in Hermes Agent, /compress <topic> to get the compaction model to retain more information on the topic you want it to keep in memory mos…

AgentsDGX agent

Nous Research has introduced a new feature in their Hermes Agent system that allows users to use the `/compress ` command to influence how the compaction model prioritizes and retains information. Thi

very memgpt / sarah wooders coded. memory isn’t a layer, it is the system. most teams think they’re choosing a model, but they’re really cho…

AgentsDGX agent

very memgpt / sarah wooders coded. memory isn’t a layer, it is the system. most teams think they’re choosing a model, but they’re really choosing where their memory lives and like ben thompson says, o

We’re thrilled to announce @MiniMax_AI M2.7 is now available Day-0 on Fireworks for commercial use. This self-evolving agentic model deliver…

AgentsDGX agent

We’re thrilled to announce @MiniMax_AI M2.7 is now available Day-0 on Fireworks for commercial use. This self-evolving agentic model delivers frontier-level performance across: → Software engineering

10 Apr 2026

Bi-level Heterogeneous Learning for Time Series Foundation Models: A Federated Learning Approach

ResearchDGX agent

arXiv:2604.06727v1 Announce Type: new Abstract: Heterogeneity in time series data is more pronounced than in vision or language, as temporal dynamics vary substantially across domains and tasks. Exist

Development of ML model for triboelectric nanogenerator based sign language detection system

ResearchDGX agent

arXiv:2604.06220v1 Announce Type: cross Abstract: Sign language recognition (SLR) is vital for bridging communication gaps between deaf and hearing communities. Vision-based approaches suffer from occ

DisCEdge: Distributed Context Management for Large Language Models at the Edge

ResearchDGX agent

arXiv:2511.22599v2 Announce Type: replace-cross Abstract: Deploying Large Language Model (LLM) services at the edge benefits latency-sensitive and privacy-aware applications. However, the stateless na

Event-Level Detection of Surgical Instrument Handovers in Videos with Interpretable Vision Models

SafetyDGX agent

arXiv:2604.07577v1 Announce Type: new Abstract: Reliable monitoring of surgical instrument exchanges is essential for maintaining procedural efficiency and patient safety in the operating room. Automa

FBS: Modeling Native Parallel Reading inside a Transformer

ResearchDGX agent

arXiv:2601.21708v2 Announce Type: replace Abstract: Large language models (LLMs) excel across many tasks, yet inference is still dominated by strictly token-by-token autoregression. Existing accelerat

I rest my case: Mythos isn’t AGI. It’s not even better at biology than the last model. It’s tuned to particular things, not a giant advance …

SafetyDGX agent

I rest my case: Mythos isn’t AGI. It’s not even better at biology than the last model. It’s tuned to particular things, not a giant advance towards general intelligence. Same as it ever was. @GaryMarc

Resource-constrained Amazons chess decision framework integrating large language models and graph attention

ResearchDGX agent

arXiv:2603.10512v2 Announce Type: replace Abstract: Artificial intelligence has advanced significantly through the development of intelligent game-playing systems, providing rigorous testbeds for deci

Toward Personalized Darts Training: A Data-Driven Framework Based on Skeleton-Based Biomechanical Analysis and Motion Modeling

Local AiDGX agent

arXiv:2604.01130v3 Announce Type: replace Abstract: As sports training becomes more data-driven, traditional dart coaching based mainly on experience and visual observation is increasingly inadequate

Towards Privacy-Preserving Large Language Model: Text-free Inference Through Alignment and Adaptation

SafetyDGX agent

arXiv:2604.06831v1 Announce Type: cross Abstract: Current LLM-based services typically require users to submit raw text regardless of its sensitivity. While intuitive, such practice introduces substan

Transforming the Voice of the Customer: Large Language Models for Identifying Customer Needs

TutorialsDGX agent

arXiv:2503.01870v2 Announce Type: replace Abstract: Identifying customer needs (CNs) is fundamental to product innovation and marketing strategy. Yet for over thirty years, Voice-of-the-Customer (VOC)

VLMShield: Efficient and Robust Defense of Vision-Language Models against Malicious Prompts

SafetyDGX agent

arXiv:2604.06502v1 Announce Type: new Abstract: Vision-Language Models (VLMs) face significant safety vulnerabilities from malicious prompt attacks due to weakened alignment during visual integration.

9 Apr 2026

(1/5) FP4 hardware is here, but 4-bit attention still kills model quality, blocking true end-to-end FP4 serving. To fix that, we propose Att…

HardwareDGX agent

(1/5) FP4 hardware is here, but 4-bit attention still kills model quality, blocking true end-to-end FP4 serving. To fix that, we propose Attn-QAT, the first systematic study of quantization-aware trai

A common theme at @aiDotEngineer @swyx @steipete 🦞 Lot’s of fun and meeting great people, my talk about model inference at @superlinked is …

ToolsDGX agent

Fardis Makraduli ([@f_makraduli](https://x.com/f_makraduli)) shared a post about attending the AI Engineer ([@aiDotEngineer](https://x.com/aiDotEngineer)) event, noting shared themes and networking...

At 15:10 today, I’ll be speaking about our SWE-rebench leaderboard at AI Engineer Europe. I'll cover how we build evals and how models cheat…

ToolsDGX agent

At 15:10 today, I’ll be speaking about our SWE-rebench leaderboard at AI Engineer Europe. I'll cover how we build evals and how models cheat! Come listen and let's chat! So far, this is the coolest ap

we got LangPod before GTA6 🙏 this series is gonna be sick - real stuff that breaks with agents, mental models, evals, tooling with some of …

ApplicationsDGX agent

we got LangPod before GTA6 🙏 this series is gonna be sick - real stuff that breaks with agents, mental models, evals, tooling with some of the best builders across industry great to openly share all t

17 Aug 2026

Consistent Model Chasing Is Minimax Optimal: The Exact Value of Scalar Adversarial Adaptive Control under Large Parametric Uncertainty

SafetyDGX agent

arXiv:2608.13651v1 Announce Type: cross Abstract: We solve exactly a fundamental problem of adaptive control against adversarial disturbances: regulate the scalar system x_{t+1} = ax_t + u_t + w_t, x_

Cross-Disciplinary Taxonomy and Modeling of Misunderstanding Generation, Amplification, and Detection, from Pragmatics to AI Agents

ResearchDGX agent

arXiv:2608.13604v1 Announce Type: new Abstract: Detection of misunderstanding is an urgent problem to solve because communication has moved away from real-time, in-person interaction and is increasing

EXL3 seems to be fading from the r/LocalLLaMa consciousness, and while I suspected it, I'm surprised at this point in time.

Model ReleasesDGX agent

EXL3 is an alternative to llama.cpp. And while there is extensive tooling for llama.cpp, EXL3's primary deployment (TabbyAPI), has a OpenAI compatible API so it shouldn't matter. Why won't this tool m

Intern-S2-Mobius: Foundation Model with Decoupled Knowledge and Reasoning

ResearchDGX agent

arXiv:2608.14290v1 Announce Type: new Abstract: We introduce Mobius-v0, an architecture that comprises a globally shared Memory (FFN) that stores knowledge vectors and multiple Reasoners (Self-Attn) t

Multiphase-Diff: Diffusion-Based Generative Modeling for High-Contrast Multiphase Physical Systems with Sharp Interfaces

ResearchDGX agent

arXiv:2608.13669v1 Announce Type: new Abstract: Physics-constrained diffusion for high-contrast, sharp-interface multiphase fields faces three coupled difficulties. At coefficient jumps, expanded poin

PRM-as-a-Judge 1.5: A Toolkit for Robot Process Assessment

Model ReleasesDGX agent

arXiv:2608.14284v1 Announce Type: cross Abstract: Fine-grained robotic evaluation matters for understanding embodied models, going beyond binary success rates and rule-based process scores. We present

Scaling Creative Writing Beyond Story-Centric Data with Attribute-Guided Genre Expansion

ResearchDGX agent

arXiv:2608.13947v1 Announce Type: new Abstract: High-quality creative writing data for large language models (LLMs) remains dominated by story-centric data, limiting models' ability to follow the stru

Secret-Stego Dissimilarity as a Design Axis: Invertible Coverless Image Steganography with Diffusion Models

ResearchDGX agent

arXiv:2608.13597v1 Announce Type: cross Abstract: Coverless image steganography (CIS) synthesizes a stego image rather than modifying an existing cover image, enabling authorized recipients to reconst

← Previous
1…224225226227228…1018
Next →