AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,587 results
Research

Speaker effects in language comprehension: An integrative model of language and speaker processing

DGX agent

arXiv:2412.07238v3 Announce Type: replace Abstract: The identity of a speaker influences language comprehension through modulating perception and expectation. This review explores speaker effects and

researcharxiv-cs-cl
15 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

TeRA: Vector-based Random Tensor Network for High-Rank Adaptation of Large Language Models

DGX agent

arXiv:2509.03234v2 Announce Type: replace Abstract: Parameter-Efficient Fine-Tuning (PEFT) methods, such as Low-Rank Adaptation (LoRA), have significantly reduced the number of trainable parameters ne

model-releasesarxiv-cs-lg
15 Apr 2026
Safety

Thinking Sparks!: Emergent Attention Heads in Reasoning Models During Post Training

DGX agent

arXiv:2509.25758v2 Announce Type: replace Abstract: The remarkable capabilities of modern large reasoning models are largely unlocked through post-training techniques such as supervised fine-tuning (S

safetyarxiv-cs-ai
15 Apr 2026
Tools

Training code and models are live on Hugging Face. Dan Fu (Together AI's VP of Kernels) led the work. Together AI provided compute. Blog: ht…

DGX agent

Training code and models are live on Hugging Face. Dan Fu (Together AI's VP of Kernels) led the work. Together AI provided compute. Blog: https://www.together.ai/blog/parcae Paper: https://arxiv.org/a

toolstogether-ai--x
15 Apr 2026
Industry

World models are soo March 2026. What's your AllBirds strategy is what matters today. How are you competing against AllBirds? Are you buildi…

DGX agent

Cristobal Valenzuela, co-founder and CEO of Runway, posted a social media comment suggesting that 'world models' — a major AI topic of early 2026 — are already becoming passé, using the footwear brand

industrycristobal-valenzuela--x
15 Apr 2026
Research

A Survey of Inductive Reasoning for Large Language Models

DGX agent

arXiv:2510.10182v2 Announce Type: replace-cross Abstract: Reasoning is an important task for large language models (LLMs). Among all the reasoning paradigms, inductive reasoning is one of the fundamen

researcharxiv-cs-ai
14 Apr 2026
Model Releases

AdaQE-CG: Adaptive Query Expansion for Web-Scale Generative AI Model and Data Card Generation

DGX agent

arXiv:2604.09617v1 Announce Type: new Abstract: Transparent and standardized documentation is essential for building trustworthy generative AI (GAI) systems. However, existing automated methods for ge

model-releasesarxiv-cs-ai
14 Apr 2026
Tutorials

Asymptotic Learning Curves for Diffusion Models with Random Features Score and Manifold Data

DGX agent

arXiv:2603.22962v2 Announce Type: replace Abstract: We study the theoretical behavior of denoising score matching--the learning task associated to diffusion models--when the data distribution is suppo

tutorialsarxiv-cs-lg
14 Apr 2026
Applications

Bootstrapping Video Semantic Segmentation Model via Distillation-assisted Test-Time Adaptation

DGX agent

arXiv:2604.10950v1 Announce Type: new Abstract: Fully supervised Video Semantic Segmentation (VSS) relies heavily on densely annotated video data, limiting practical applicability. Alternatively, appl

applicationsarxiv-cs-cv
14 Apr 2026
Research

Breaking the KV Cache Bottleneck: Fan Duality Model Achieves O(1) Decode Memory with Superior Associative Recall

DGX agent

arXiv:2604.07716v2 Announce Type: replace Abstract: We present FDM (Fan Duality Model), a linear sequence architecture that resolves the fundamental tension between memory efficiency and associative r

researcharxiv-cs-lg
14 Apr 2026
Research

Cognitive Training for Language Models: Towards General Capabilities via Cross-Entropy Games

DGX agent

arXiv:2603.22479v3 Announce Type: replace-cross Abstract: Defining a constructive process to build general capabilities for language models in an automatic manner is considered an open problem in arti

researcharxiv-cs-ai
14 Apr 2026
Research

Computational Implementation of a Model of Category-Theoretic Metaphor Comprehension

DGX agent

arXiv:2604.10035v1 Announce Type: cross Abstract: In this study, we developed a computational implementation for a model of metaphor comprehension based on the theory of indeterminate natural transfor

researcharxiv-cs-ai
14 Apr 2026
Model Releases

CounterBench: Evaluating and Improving Counterfactual Reasoning in Large Language Models

DGX agent

arXiv:2502.11008v2 Announce Type: replace Abstract: Counterfactual reasoning is widely recognized as one of the most challenging and intricate aspects of causality in artificial intelligence. In this

model-releasesarxiv-cs-cl
14 Apr 2026
Research

Critical-CoT: A Robust Defense Framework against Reasoning-Level Backdoor Attacks in Large Language Models

DGX agent

arXiv:2604.10681v1 Announce Type: cross Abstract: Large Language Models (LLMs), despite their impressive capabilities across domains, have been shown to be vulnerable to backdoor attacks. Prior backdo

researcharxiv-cs-ai
14 Apr 2026
Safety

CROP: Conservative Reward for Model-based Offline Policy Optimization

DGX agent

arXiv:2310.17245v2 Announce Type: replace-cross Abstract: Offline reinforcement learning (RL) aims to optimize a policy using collected data without online interactions. Model-based approaches are par

safetyarxiv-cs-ai
14 Apr 2026
Hardware

Deep Optimizer States: Towards Scalable Training of Transformer Models Using Interleaved Offloading

DGX agent

arXiv:2410.21316v2 Announce Type: replace-cross Abstract: Transformers and large language models~(LLMs) have seen rapid adoption in all domains. Their sizes have exploded to hundreds of billions of pa

hardwarearxiv-cs-ai
14 Apr 2026
Research

Delving Aleatoric Uncertainty in Medical Image Segmentation via Vision Foundation Models

DGX agent

arXiv:2604.10963v1 Announce Type: new Abstract: Medical image segmentation supports clinical workflows by precisely delineating anatomical structures and lesions. However, medical image datasets medic

researcharxiv-cs-ai
14 Apr 2026
Local Ai

Different types of syntactic agreement recruit the same units within large language models

DGX agent

arXiv:2512.03676v2 Announce Type: replace Abstract: Large language models (LLMs) can reliably distinguish grammatical from ungrammatical sentences, but how grammatical knowledge is represented within

local-aiarxiv-cs-cl
14 Apr 2026
Local Ai

Do vision models perceive illusory motion in static images like humans?

DGX agent

arXiv:2604.09853v1 Announce Type: new Abstract: Understanding human motion processing is essential for building reliable, human-centered computer vision systems. Although deep neural networks (DNNs) a

local-aiarxiv-cs-cv
14 Apr 2026
Research

Efficient Process Reward Modeling via Contrastive Mutual Information

DGX agent

arXiv:2604.10660v1 Announce Type: cross Abstract: Recent research has devoted considerable effort to verifying the intermediate reasoning steps of chain-of-thought (CoT) trajectories using process rew

researcharxiv-cs-ai
14 Apr 2026
Safety

Efficient Training for Cross-lingual Speech Language Models

DGX agent

arXiv:2604.11096v1 Announce Type: cross Abstract: Currently, large language models (LLMs) predominantly focus on the text modality. To enable more natural human-AI interaction, speech LLMs are emergin

safetyarxiv-cs-ai
14 Apr 2026
Safety

Empowering Video Translation using Multimodal Large Language Models

DGX agent

arXiv:2604.11283v1 Announce Type: new Abstract: Recent developments in video translation have further enhanced cross-lingual access to video content, with multimodal large language models (MLLMs) play

safetyarxiv-cs-cv
14 Apr 2026
Tutorials

Energy-oriented Diffusion Bridge for Image Restoration with Foundational Diffusion Models

DGX agent

arXiv:2604.10983v1 Announce Type: new Abstract: Diffusion bridge models have shown great promise in image restoration by explicitly connecting clean and degraded image distributions. However, they oft

tutorialsarxiv-cs-cv
14 Apr 2026
Tutorials

EviCare: Enhancing Diagnosis Prediction with Deep Model-Guided Evidence for In-Context Reasoning

DGX agent

arXiv:2604.10455v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have enabled promising progress in diagnosis prediction from electronic health records (EHRs). However,

tutorialsarxiv-cs-cl
14 Apr 2026
Safety

Fake-HR1: Rethinking Reasoning of Vision Language Model for Synthetic Image Detection

DGX agent

arXiv:2602.10042v3 Announce Type: replace-cross Abstract: Recent studies have demonstrated that incorporating Chain-of-Thought (CoT) reasoning into the detection process can enhance a model's ability

safetyarxiv-cs-ai
14 Apr 2026
Applications

Generating Multiple-Choice Knowledge Questions with Interpretable Difficulty Estimation using Knowledge Graphs and Large Language Models

DGX agent

arXiv:2604.10748v1 Announce Type: cross Abstract: Generating multiple-choice questions (MCQs) with difficulty estimation remains challenging in automated MCQ-generation systems used in adaptive, AI-as

applicationsarxiv-cs-ai
14 Apr 2026
Research

HFI: A unified framework for training-free detection and implicit watermarking of latent diffusion model generated images

DGX agent

arXiv:2412.20704v2 Announce Type: replace Abstract: Dramatic advances in the quality of the latent diffusion models (LDMs) also led to the malicious use of AI-generated images. While current AI-genera

researcharxiv-cs-cv
14 Apr 2026
Research

HOG-Layout: Hierarchical 3D Scene Generation, Optimization and Editing via Vision-Language Models

DGX agent

arXiv:2604.10772v1 Announce Type: new Abstract: 3D layout generation and editing play a crucial role in Embodied AI and immersive VR interaction. However, manual creation requires tedious labor, while

researcharxiv-cs-cv
14 Apr 2026
Agents

Human Centered Non Intrusive Driver State Modeling Using Personalized Physiological Signals in Real World Automated Driving

DGX agent

arXiv:2604.11549v1 Announce Type: cross Abstract: In vehicles with partial or conditional driving automation (SAE Levels 2-3), the driver remains responsible for supervising the system and responding

agentsarxiv-cs-lg
14 Apr 2026
Research

Human-like Working Memory Interference in Large Language Models

DGX agent

arXiv:2604.09670v1 Announce Type: cross Abstract: Intelligent systems must maintain and manipulate task-relevant information online to adapt to dynamic environments and changing goals. This capacity,

researcharxiv-cs-ai
14 Apr 2026
Model Releases

I have a Macbook AIR M5 Base and I want to run an Agentic Coding program, similar to Claude Code or Codex. Besides the model, how do I do it? I've already tried with Ollama, VS Code, Opencode, and haven't been able to. (I'm not a developer, sorry)

DGX agent

This Reddit thread addresses a common challenge for non-developers trying to run a local agentic coding assistant on a MacBook Air M5: while tools like Ollama, VS Code, and OpenCode are the right piec

model-releasesr-ollama
14 Apr 2026
Research

Improving Pediatric Emergency Department Triage with Modality Dropout in Late Fusion Multimodal EHR Models

DGX agent

arXiv:2604.09905v1 Announce Type: new Abstract: Emergency department triage relies heavily on both quantitative vital signs and qualitative clinical notes, yet multimodal machine learning models predi

researcharxiv-cs-lg
14 Apr 2026
Model Releases

Knowledge Integration in Differentiable Models: A Comparative Study of Data-Driven, Soft-Constrained, and Hard-Constrained Paradigms for Identification and Control of the Single Machine Infinite Bus System

DGX agent

arXiv:2602.09667v2 Announce Type: replace Abstract: Integrating domain knowledge into neural networks is a central challenge in scientific machine learning. Three paradigms have emerged -- data-driven

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

Large Language Models Can Help Mitigate Barren Plateaus in Quantum Neural Networks

DGX agent

arXiv:2502.13166v3 Announce Type: replace-cross Abstract: In the era of noisy intermediate-scale quantum (NISQ) computing, Quantum Neural Networks (QNNs) have emerged as a promising approach for vario

model-releasesarxiv-cs-ai
14 Apr 2026
Local Ai

Machine-learning modeling of magnetization dynamics in quasi-equilibrium and driven metallic spin systems

DGX agent

arXiv:2604.11513v1 Announce Type: cross Abstract: We review recent advances in machine-learning (ML) force-field methods for large-scale Landau-Lifshitz-Gilbert (LLG) simulations of metallic spin syst

local-aiarxiv-cs-lg
14 Apr 2026
Model Releases

MedVeriSeg: Teaching MLLM-Based Medical Segmentation Models to Verify Query Validity Without Extra Training

DGX agent

arXiv:2604.10242v1 Announce Type: new Abstract: Despite recent advances in MLLM-based medical image segmentation, existing LISA-like methods cannot reliably reject false queries and often produce hall

model-releasesarxiv-cs-cv
14 Apr 2026
Research

Merging Triggers, Breaking Backdoors: Defensive Poisoning for Instruction-Tuned Language Models

DGX agent

arXiv:2601.04448v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have greatly advanced Natural Language Processing (NLP), particularly through instruction tuning, which enables b

researcharxiv-cs-ai
14 Apr 2026
Research

NeuroFlow: Toward Unified Visual Encoding and Decoding from Neural Activity

DGX agent

arXiv:2604.09817v1 Announce Type: new Abstract: Visual encoding and decoding models act as gateways to understanding the neural mechanisms underlying human visual perception. Typically, visual encodin

researcharxiv-cs-lg
14 Apr 2026
Local Ai

New LTX model soon

DGX agent

This r/StableDiffusion post likely discussed the anticipated release of an upcoming LTX video generation model from Lightricks, previewing improvements over existing versions before a formal announcem

local-air-stablediffusion
14 Apr 2026
Local Ai

Optimizing Large Language Models: Metrics, Energy Efficiency, and Case Study Insights

DGX agent

arXiv:2504.06307v2 Announce Type: replace-cross Abstract: The rapid adoption of large language models (LLMs) has led to significant energy consumption and carbon emissions, posing a critical challenge

local-aiarxiv-cs-ai
14 Apr 2026
Applications

PnP-CM: Consistency Models as Plug-and-Play Priors for Inverse Problems

DGX agent

arXiv:2509.22736v2 Announce Type: replace-cross Abstract: Diffusion models have found extensive use in solving inverse problems, by sampling from an approximate posterior distribution of data given th

applicationsarxiv-cs-ai
14 Apr 2026
Tutorials

Quantum-Gated Task-interaction Knowledge Distillation for Pre-trained Model-based Class-Incremental Learning

DGX agent

arXiv:2604.11112v1 Announce Type: cross Abstract: Class-incremental learning (CIL) aims to continuously accumulate knowledge from a stream of tasks and construct a unified classifier over all seen cla

tutorialsarxiv-cs-cv
14 Apr 2026
Industry

Randomly made the HN frontpage with my latest blog post on function calling and open source models. https://www.thetypicalset.com/blog/gramm…

DGX agent

Remi Louf shared that a blog post he wrote about function calling and open source models unexpectedly reached the Hacker News frontpage. The post, hosted on thetypicalset.com, likely explores how open

industryclem-delangue--x
14 Apr 2026
Tutorials

Ro-SLM: Onboard Small Language Models for Robot Task Planning and Operation Code Generation

DGX agent

arXiv:2604.10929v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) provide robots with contextual reasoning abilities to comprehend human instructions. Yet, current LLM-en

tutorialsarxiv-cs-ro
14 Apr 2026
Model Releases

Scone: Bridging Composition and Distinction in Subject-Driven Image Generation via Unified Understanding-Generation Modeling

DGX agent

arXiv:2512.12675v2 Announce Type: replace-cross Abstract: Subject-driven image generation has advanced from single- to multi-subject composition, while neglecting distinction, the ability to distingui

model-releasesarxiv-cs-ai
14 Apr 2026
Research

STU-PID: Steering Token Usage via PID Controller for Efficient Large Language Model Reasoning

DGX agent

arXiv:2506.18831v2 Announce Type: replace Abstract: Large Language Models employing extended chain-of-thought (CoT) reasoning often suffer from the overthinking phenomenon, generating excessive and re

researcharxiv-cs-cl
14 Apr 2026
Model Releases

TempusBench: An Evaluation Framework for Time-Series Forecasting

DGX agent

arXiv:2604.11529v1 Announce Type: new Abstract: Foundation models have transformed natural language processing and computer vision, and a rapidly growing literature on time-series foundation models (T

model-releasesarxiv-cs-lg
14 Apr 2026
Tutorials

Transformers Learn Latent Mixture Models In-Context via Mirror Descent

DGX agent

arXiv:2604.10848v1 Announce Type: new Abstract: Sequence modelling requires determining which past tokens are causally relevant from the context and their importance: a process inherent to the attenti

tutorialsarxiv-cs-lg
14 Apr 2026
← Previous
1…180181182183184…1263
Next →