AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,001 results
Research

AFL: A Single-Round Analytic Approach for Federated Learning with Pre-trained Models

DGX agent

arXiv:2405.16240v3 Announce Type: replace Abstract: In this paper, we introduce analytic federated learning (AFL), a new training paradigm that brings analytical (i.e., closed-form) solutions to the f

researcharxiv-cs-lg
10 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Industry

Apple's head of cloud says Open Source models will address 90% of the use case

DGX agent

I was unable to retrieve the specific Reddit thread or locate reliable sourced details about Apple's head of cloud making a statement that open source models will address 90% of use cases. The sear...

industryr-chatgpt
10 Apr 2026
Tutorials

Attribution-Driven Explainable Intrusion Detection with Encoder-Based Large Language Models

DGX agent

arXiv:2604.06266v1 Announce Type: cross Abstract: Software-Defined Networking (SDN) improves network flexibility but also increases the need for reliable and interpretable intrusion detection. Large L

tutorialsarxiv-cs-ai
10 Apr 2026
Research

Beyond the Mean: Modelling Annotation Distributions in Continuous Affect Prediction

DGX agent

arXiv:2604.07198v1 Announce Type: new Abstract: Emotion annotation is inherently subjective and cognitively demanding, producing signals that reflect diverse perceptions across annotators rather than

researcharxiv-cs-lg
10 Apr 2026
Model Releases

DISSECT: Diagnosing Where Vision Ends and Language Priors Begin in Scientific VLMs

DGX agent

arXiv:2604.06250v1 Announce Type: cross Abstract: When asked to describe a molecular diagram, a Vision-Language Model correctly identifies ``a benzene ring with an -OH group.'' When asked to reason ab

model-releasesarxiv-cs-ai
10 Apr 2026
Research

'Don't Do That!': Guiding Embodied Systems through Large Language Model-based Constraint Generation

DGX agent

arXiv:2506.04500v3 Announce Type: replace-cross Abstract: Recent advancements in large language models (LLMs) have spurred interest in robotic navigation that incorporates complex spatial, mathematica

researcharxiv-cs-ro
10 Apr 2026
Research

Energy-Regularized Spatial Masking: A Novel Approach to Enhancing Robustness and Interpretability in Vision Models

DGX agent

arXiv:2604.06893v1 Announce Type: cross Abstract: Deep convolutional neural networks achieve remarkable performance by exhaustively processing dense spatial feature maps, yet this brute-force strategy

researcharxiv-cs-lg
10 Apr 2026
Model Releases

Flow Motion Policy: Manipulator Motion Planning with Flow Matching Models

DGX agent

arXiv:2604.07084v1 Announce Type: cross Abstract: Open-loop end-to-end neural motion planners have recently been proposed to improve motion planning for robotic manipulators. These methods enable plan

model-releasesarxiv-cs-ai
10 Apr 2026
Research

Hallucination as output-boundary misclassification: a composite abstention architecture for language models

DGX agent

arXiv:2604.06195v1 Announce Type: cross Abstract: Large language models often produce unsupported claims. We frame this as a misclassification error at the output boundary, where internally generated

researcharxiv-cs-ai
10 Apr 2026
Research

Hierarchical Feature Learning for Medical Point Clouds via State Space Model

DGX agent

arXiv:2504.13015v3 Announce Type: replace Abstract: Deep learning-based point cloud modeling has been widely investigated as an indispensable component of general shape analysis. Recently, transformer

researcharxiv-cs-cv
10 Apr 2026
Research

Incentive-Aware Multi-Fidelity Optimization for Generative Advertising in Large Language Models

DGX agent

arXiv:2604.06263v1 Announce Type: cross Abstract: Generative advertising in large language model (LLM) responses requires optimizing sponsorship configurations under two strict constraints: the strate

researcharxiv-cs-ai
10 Apr 2026
Tutorials

Learning is Forgetting: LLM Training As Lossy Compression

DGX agent

arXiv:2604.07569v1 Announce Type: cross Abstract: Despite the increasing prevalence of large language models (LLMs), we still have a limited understanding of how their representational spaces are stru

tutorialsarxiv-cs-cl
10 Apr 2026
Safety

LINE: LLM-based Iterative Neuron Explanations for Vision Models

DGX agent

arXiv:2604.08039v1 Announce Type: new Abstract: Interpreting the concepts encoded by individual neurons in deep neural networks is a crucial step towards understanding their complex decision-making pr

safetyarxiv-cs-cv
10 Apr 2026
Model Releases

LongWriter-Zero: Mastering Ultra-Long Text Generation via Reinforcement Learning

DGX agent

arXiv:2506.18841v3 Announce Type: replace-cross Abstract: Ultra-long generation by large language models (LLMs) is a widely demanded scenario, yet it remains a significant challenge due to their maxim

model-releasesarxiv-cs-ai
10 Apr 2026
Tutorials

ODE-free Neural Flow Matching for One-Step Generative Modeling

DGX agent

arXiv:2604.06413v1 Announce Type: new Abstract: Diffusion and flow matching models generate samples by learning time-dependent vector fields whose integration transports noise to data, requiring tens

tutorialsarxiv-cs-lg
10 Apr 2026
Agents

One Life to Learn: Inferring Symbolic World Models for Stochastic Environments from Unguided Exploration

DGX agent

arXiv:2510.12088v2 Announce Type: replace Abstract: Symbolic world modeling requires inferring and representing an environment's transitional dynamics as an executable program. Prior work has focused

agentsarxiv-cs-ai
10 Apr 2026
Model Releases

Rethinking Entropy Allocation in LLM-based ASR: Understanding the Dynamics between Speech Encoders and LLMs

DGX agent

arXiv:2604.08003v1 Announce Type: cross Abstract: Integrating large language models (LLMs) into automatic speech recognition (ASR) has become a dominant paradigm. Although recent LLM-based ASR models

model-releasesarxiv-cs-cl
10 Apr 2026
Research

STQuant: Spatio-Temporal Adaptive Framework for Optimizer Quantization in Large Multimodal Model Training

DGX agent

arXiv:2604.06836v2 Announce Type: new Abstract: Quantization is an effective way to reduce the memory cost of large-scale model training. However, most existing methods adopt fixed-precision policies,

researcharxiv-cs-lg
10 Apr 2026
Tutorials

The quality of talks at @aiDotEngineer is insane, being able to learn about diffusion models and flow mapping from @GoogleDeepMind’s @sediel…

DGX agent

The AI Engineer Summit (@aiDotEngineer) is a highly regarded technical conference featuring speakers from leading AI organizations including Google DeepMind, Anthropic, and OpenAI, known for its de...

tutorialsswyx--x
10 Apr 2026
Model Releases

Visual prompting reimagined: The power of the Activation Prompts

DGX agent

arXiv:2604.06440v1 Announce Type: cross Abstract: Visual prompting (VP) has emerged as a popular method to repurpose pretrained vision models for adaptation to downstream tasks. Unlike conventional mo

model-releasesarxiv-cs-lg
10 Apr 2026
Agents

a useful mental model on how teams can think about good data design to improve their models/agents: Evals ~= Training Data ~= Environments -…

DGX agent

a useful mental model on how teams can think about good data design to improve their models/agents: Evals ~= Training Data ~= Environments - in Classical Deep Learning, we learn from each training exa

agentsharrison-chase--x
9 Apr 2026
Agents

Another banger paper from Microsoft. Why it's a big deal: It teaches reasoning models to compress their own chain-of-thought mid-generation.…

DGX agent

Another banger paper from Microsoft. Why it's a big deal: It teaches reasoning models to compress their own chain-of-thought mid-generation. The most interesting finding isn't the 2-3x memory savings

agentsdair-ai--x
9 Apr 2026
Agents

This 10-min read from @Vtrivedy10 changes how you build AI agents. Most people are stuck in the same loop; switching models when agents brea…

DGX agent

This 10-min read from @Vtrivedy10 changes how you build AI agents. Most people are stuck in the same loop; switching models when agents break. The reframe: evals are the training data for your harness

agentsharrison-chase--x
9 Apr 2026
Industry

Anthropic had the most powerful cyber-security model in the history of this world and their internal code based still leaked? We should assu…

DGX agent

Anthropic had the most powerful cyber-security model in the history of this world and their internal code based still leaked? We should assume everyone can be compromised, and build systems that keep

industryclem-delangue--x
8 Apr 2026
Agents

GLM-5.1 gives teams a stronger model for coding, tool use, and sustained agent performance on Together AI. Learn more: http://www.together.a…

DGX agent

GLM-5.1 is Z.ai's post-training upgrade to GLM-5, now available on Together AI, delivering a 28% coding performance improvement through a refined reinforcement learning pipeline while retaining the...

agentstogether-ai--x
8 Apr 2026
Agents

> which is exactly why we believe memory should live outside of model providers open harness = open memory which everyone should want!

DGX agent

> which is exactly why we believe memory should live outside of model providers open harness = open memory which everyone should want! The new Anthropic managed agents API is basically the Letta API t

agentsharrison-chase--x
8 Apr 2026
Model Releases

Introducing GLM-5.1 for understanding research papers 🚀 Highlight any section of a paper to ask questions and “@” other papers for quick co…

DGX agent

AlphaXiv introduced GLM-5.1 as the underlying model powering its research paper understanding features on the alphaXiv platform, enabling users to highlight any section of a paper to ask contextual...

model-releaseszhipu-ai--x
7 Apr 2026
Research

A Survey of Large Models in Sports

DGX agent

arXiv:2608.14377v1 Announce Type: new Abstract: Sports have witnessed growing global enthusiasm in recent years, serving as a vital force for physical health, cultural exchange, social connection, and

researcharxiv-cs-cl
17 Aug 2026
Research

Classical Limits of Spectral Filtering in Quantum Generative Models

DGX agent

arXiv:2608.14169v1 Announce Type: cross Abstract: Spectral filtering has been proposed as a route to regularization in quantum generative models: the quantum Fourier transform exposes the amplitude sp

researcharxiv-cs-lg
17 Aug 2026
Model Releases

CPI-Bench: A Comprehensive,Practical and Intelligent Benchmark for Real-World Image Editing

DGX agent

arXiv:2608.14546v1 Announce Type: new Abstract: With the rapid advancement of image editing models and their widespread application across various domains, there is an increasingly urgent need to depl

model-releasesarxiv-cs-cv
17 Aug 2026
Tutorials

CytoBERT: A Foundation Model for Cytometry Data

DGX agent

arXiv:2608.14414v1 Announce Type: new Abstract: Cytometry measures the complex characteristics of single cells (e.g., counts and protein expression of immune cells) and is widely used across immunolog

tutorialsarxiv-cs-lg
17 Aug 2026
Model Releases

Exponential-Family Membership Inference: From LiRA and RMIA to BaVarIA

DGX agent

arXiv:2603.11799v2 Announce Type: replace Abstract: Membership inference attacks (MIAs) are becoming standard tools for auditing the privacy of machine learning models. The leading attacks -- LiRA (Ca

model-releasesarxiv-cs-lg
17 Aug 2026
Research

From crown candidates to neighborhood screening: integrating optical GeoAI and spatial modeling for urban-canopy assessment in Davis, California

DGX agent

arXiv:2608.13856v1 Announce Type: cross Abstract: Timely urban-canopy information is essential for linking remote sensing with heat, mobility, and neighborhood planning. We developed an optical GeoAI

researcharxiv-cs-cv
17 Aug 2026
Safety

Understanding and Mitigating Over-refusal for Large Language Models via Representation Intervention

DGX agent

arXiv:2511.19009v2 Announce Type: replace-cross Abstract: Large language models (LLMs) demonstrate powerful capabilities across various natural language processing tasks,yet their inherent safety vuln

safetyarxiv-cs-cl
17 Aug 2026
Research

8/19, join us for the #ACMTechTalk, 'From Conventional LLMs to Reasoning Models to Agents,' w/AI & LLM Research Engineer @rasbt. ACM Practit…

DGX agent

8/19, join us for the #ACMTechTalk, 'From Conventional LLMs to Reasoning Models to Agents,' w/AI & LLM Research Engineer @rasbt. ACM Practitioner Board Co-Chaior @marlene_zw (@Microsoft) will moderate

researchsebastian-raschka--x
14 Aug 2026
Model Releases

A Probe Direction Is a Property of Its Prompt

DGX agent

arXiv:2608.13329v1 Announce Type: new Abstract: A model that behaves differently when it senses it is being tested would undermine the evaluations we rely on, so recent work has sought to read that se

model-releasesarxiv-cs-lg
14 Aug 2026
Research

Capstan-driven Continuum Surgical Robot: Design, Modeling, and Perception

DGX agent

arXiv:2608.13396v1 Announce Type: new Abstract: Shape and force sensing have long been critical bottlenecks in the development of compact capstan-driven continuum surgical robots, primarily due to the

researcharxiv-cs-ro
14 Aug 2026
Research

Demand Transfer Estimation at Scale via Restricted Logit Modeling

DGX agent

arXiv:2608.12680v1 Announce Type: cross Abstract: Item demand forecasting is an integral component of store assortment optimization. Existing literature focuses on learning a suitable customer choice

researcharxiv-cs-ai
14 Aug 2026
Safety

DiffGRM: Diffusion-based Generative Recommendation Model

DGX agent

arXiv:2510.21805v2 Announce Type: replace-cross Abstract: Generative recommendation (GR) is an emerging paradigm that represents each item via a tokenizer as an n-digit semantic ID (SID) and predicts

safetyarxiv-cs-ai
14 Aug 2026
Research

From Observation to Intervention: Memory in Brains and Large Language Models

DGX agent

arXiv:2608.12377v1 Announce Type: cross Abstract: Brains and large language models (LLMs) are fundamentally different memory systems, but they can be compared through shared functional questions: wher

researcharxiv-cs-ai
14 Aug 2026
Safety

Intern-S2-Preview: Scientific Agentic Foundation Model

DGX agent

arXiv:2608.13505v1 Announce Type: cross Abstract: Scientific discovery increasingly requires AI systems that can reason over scientific evidence of heterogeneous modalities, interact with scientific t

safetyarxiv-cs-cl
14 Aug 2026
Safety

JailWAM: Jailbreaking World Action Models in Robot Control

DGX agent

arXiv:2604.05498v2 Announce Type: replace Abstract: World Action Models (WAMs) have emerged as a promising paradigm for robotic manipulation, enabling physical interaction across diverse tasks and env

safetyarxiv-cs-ro
14 Aug 2026
Model Releases

Jeremy's excellent work here is a great illustration of a very powerful type of approach: LLM-guided on-the-fly synthesis of a symbolic worl…

DGX agent

Jeremy's excellent work here is a great illustration of a very powerful type of approach: LLM-guided on-the-fly synthesis of a symbolic world model, i.e. making sense of the world by writing executabl

model-releasesfrancois-chollet--x
14 Aug 2026
Hardware

Qwen3.8-Max is live on DigitalOcean Serverless Inference. Launching side by side with DigitalOcean as our Day 0 launch partner. Big model. S…

DGX agent

Qwen3.8-Max is live on DigitalOcean Serverless Inference. Launching side by side with DigitalOcean as our Day 0 launch partner. Big model. Smooth sailing. Now on DigitalOcean.🌊🏄‍♀️ @digitalocean Now a

hardwareqwen--x
14 Aug 2026
Local Ai

Request: More Transparency on Ollama Cloud Subscriptions

DGX agent

I've been an Ollama Cloud subscriber for ~6 months. I generally use the the latest GLM models available for coding as well as a personal instance of Open WebUI. I have tried to get answers directly vi

local-air-ollama
14 Aug 2026
Industry

Sources: Apple trained a China-specific LLM with Alibaba's support, which would make Apple the first foreign company to offer a proprietary AI model in China (Reuters)

DGX agent

Reuters: Sources: Apple trained a China-specific LLM with Alibaba's support, which would make Apple the first foreign company to offer a proprietary AI model in China — Apple (AAPL.O) has trained a la

industrytechmeme
14 Aug 2026
Research

Structure-aware Riemannian Growth Fields for 4D Plant Modeling

DGX agent

arXiv:2608.13007v1 Announce Type: new Abstract: In this paper, we introduce a novel framework for 4D plant growth modeling that reconstructs the continuous geometric and topological evolution of plant

researcharxiv-cs-cv
14 Aug 2026
Research

A Reality Check of Language Models as Formalizers on Constraint Satisfaction Problems

DGX agent

arXiv:2505.13252v5 Announce Type: replace Abstract: Recent work shows superior performance when using large language models (LLMs) as formalizers instead of as end-to-end solvers for symbolic reasonin

researcharxiv-cs-cl
13 Aug 2026
← Previous
1…211212213214215…1271
Next →