AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,553 results
14 Apr 2026

Different types of syntactic agreement recruit the same units within large language models

Local AiDGX agent

arXiv:2512.03676v2 Announce Type: replace Abstract: Large language models (LLMs) can reliably distinguish grammatical from ungrammatical sentences, but how grammatical knowledge is represented within

Do vision models perceive illusory motion in static images like humans?

Local AiDGX agent

arXiv:2604.09853v1 Announce Type: new Abstract: Understanding human motion processing is essential for building reliable, human-centered computer vision systems. Although deep neural networks (DNNs) a

Efficient Process Reward Modeling via Contrastive Mutual Information

ResearchDGX agent

arXiv:2604.10660v1 Announce Type: cross Abstract: Recent research has devoted considerable effort to verifying the intermediate reasoning steps of chain-of-thought (CoT) trajectories using process rew

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Efficient Training for Cross-lingual Speech Language Models

SafetyDGX agent

arXiv:2604.11096v1 Announce Type: cross Abstract: Currently, large language models (LLMs) predominantly focus on the text modality. To enable more natural human-AI interaction, speech LLMs are emergin

Empowering Video Translation using Multimodal Large Language Models

SafetyDGX agent

arXiv:2604.11283v1 Announce Type: new Abstract: Recent developments in video translation have further enhanced cross-lingual access to video content, with multimodal large language models (MLLMs) play

Energy-oriented Diffusion Bridge for Image Restoration with Foundational Diffusion Models

TutorialsDGX agent

arXiv:2604.10983v1 Announce Type: new Abstract: Diffusion bridge models have shown great promise in image restoration by explicitly connecting clean and degraded image distributions. However, they oft

EviCare: Enhancing Diagnosis Prediction with Deep Model-Guided Evidence for In-Context Reasoning

TutorialsDGX agent

arXiv:2604.10455v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have enabled promising progress in diagnosis prediction from electronic health records (EHRs). However,

Fake-HR1: Rethinking Reasoning of Vision Language Model for Synthetic Image Detection

SafetyDGX agent

arXiv:2602.10042v3 Announce Type: replace-cross Abstract: Recent studies have demonstrated that incorporating Chain-of-Thought (CoT) reasoning into the detection process can enhance a model's ability

Generating Multiple-Choice Knowledge Questions with Interpretable Difficulty Estimation using Knowledge Graphs and Large Language Models

ApplicationsDGX agent

arXiv:2604.10748v1 Announce Type: cross Abstract: Generating multiple-choice questions (MCQs) with difficulty estimation remains challenging in automated MCQ-generation systems used in adaptive, AI-as

HFI: A unified framework for training-free detection and implicit watermarking of latent diffusion model generated images

ResearchDGX agent

arXiv:2412.20704v2 Announce Type: replace Abstract: Dramatic advances in the quality of the latent diffusion models (LDMs) also led to the malicious use of AI-generated images. While current AI-genera

HOG-Layout: Hierarchical 3D Scene Generation, Optimization and Editing via Vision-Language Models

ResearchDGX agent

arXiv:2604.10772v1 Announce Type: new Abstract: 3D layout generation and editing play a crucial role in Embodied AI and immersive VR interaction. However, manual creation requires tedious labor, while

Human Centered Non Intrusive Driver State Modeling Using Personalized Physiological Signals in Real World Automated Driving

AgentsDGX agent

arXiv:2604.11549v1 Announce Type: cross Abstract: In vehicles with partial or conditional driving automation (SAE Levels 2-3), the driver remains responsible for supervising the system and responding

Human-like Working Memory Interference in Large Language Models

ResearchDGX agent

arXiv:2604.09670v1 Announce Type: cross Abstract: Intelligent systems must maintain and manipulate task-relevant information online to adapt to dynamic environments and changing goals. This capacity,

I have a Macbook AIR M5 Base and I want to run an Agentic Coding program, similar to Claude Code or Codex. Besides the model, how do I do it? I've already tried with Ollama, VS Code, Opencode, and haven't been able to. (I'm not a developer, sorry)

Model ReleasesDGX agent

This Reddit thread addresses a common challenge for non-developers trying to run a local agentic coding assistant on a MacBook Air M5: while tools like Ollama, VS Code, and OpenCode are the right piec

Improving Pediatric Emergency Department Triage with Modality Dropout in Late Fusion Multimodal EHR Models

ResearchDGX agent

arXiv:2604.09905v1 Announce Type: new Abstract: Emergency department triage relies heavily on both quantitative vital signs and qualitative clinical notes, yet multimodal machine learning models predi

Knowledge Integration in Differentiable Models: A Comparative Study of Data-Driven, Soft-Constrained, and Hard-Constrained Paradigms for Identification and Control of the Single Machine Infinite Bus System

Model ReleasesDGX agent

arXiv:2602.09667v2 Announce Type: replace Abstract: Integrating domain knowledge into neural networks is a central challenge in scientific machine learning. Three paradigms have emerged -- data-driven

Large Language Models Can Help Mitigate Barren Plateaus in Quantum Neural Networks

Model ReleasesDGX agent

arXiv:2502.13166v3 Announce Type: replace-cross Abstract: In the era of noisy intermediate-scale quantum (NISQ) computing, Quantum Neural Networks (QNNs) have emerged as a promising approach for vario

Machine-learning modeling of magnetization dynamics in quasi-equilibrium and driven metallic spin systems

Local AiDGX agent

arXiv:2604.11513v1 Announce Type: cross Abstract: We review recent advances in machine-learning (ML) force-field methods for large-scale Landau-Lifshitz-Gilbert (LLG) simulations of metallic spin syst

MedVeriSeg: Teaching MLLM-Based Medical Segmentation Models to Verify Query Validity Without Extra Training

Model ReleasesDGX agent

arXiv:2604.10242v1 Announce Type: new Abstract: Despite recent advances in MLLM-based medical image segmentation, existing LISA-like methods cannot reliably reject false queries and often produce hall

Merging Triggers, Breaking Backdoors: Defensive Poisoning for Instruction-Tuned Language Models

ResearchDGX agent

arXiv:2601.04448v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have greatly advanced Natural Language Processing (NLP), particularly through instruction tuning, which enables b

NeuroFlow: Toward Unified Visual Encoding and Decoding from Neural Activity

ResearchDGX agent

arXiv:2604.09817v1 Announce Type: new Abstract: Visual encoding and decoding models act as gateways to understanding the neural mechanisms underlying human visual perception. Typically, visual encodin

New LTX model soon

Local AiDGX agent

This r/StableDiffusion post likely discussed the anticipated release of an upcoming LTX video generation model from Lightricks, previewing improvements over existing versions before a formal announcem

Optimizing Large Language Models: Metrics, Energy Efficiency, and Case Study Insights

Local AiDGX agent

arXiv:2504.06307v2 Announce Type: replace-cross Abstract: The rapid adoption of large language models (LLMs) has led to significant energy consumption and carbon emissions, posing a critical challenge

PnP-CM: Consistency Models as Plug-and-Play Priors for Inverse Problems

ApplicationsDGX agent

arXiv:2509.22736v2 Announce Type: replace-cross Abstract: Diffusion models have found extensive use in solving inverse problems, by sampling from an approximate posterior distribution of data given th

Quantum-Gated Task-interaction Knowledge Distillation for Pre-trained Model-based Class-Incremental Learning

TutorialsDGX agent

arXiv:2604.11112v1 Announce Type: cross Abstract: Class-incremental learning (CIL) aims to continuously accumulate knowledge from a stream of tasks and construct a unified classifier over all seen cla

Randomly made the HN frontpage with my latest blog post on function calling and open source models. https://www.thetypicalset.com/blog/gramm…

IndustryDGX agent

Remi Louf shared that a blog post he wrote about function calling and open source models unexpectedly reached the Hacker News frontpage. The post, hosted on thetypicalset.com, likely explores how open

Ro-SLM: Onboard Small Language Models for Robot Task Planning and Operation Code Generation

TutorialsDGX agent

arXiv:2604.10929v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) provide robots with contextual reasoning abilities to comprehend human instructions. Yet, current LLM-en

Scone: Bridging Composition and Distinction in Subject-Driven Image Generation via Unified Understanding-Generation Modeling

Model ReleasesDGX agent

arXiv:2512.12675v2 Announce Type: replace-cross Abstract: Subject-driven image generation has advanced from single- to multi-subject composition, while neglecting distinction, the ability to distingui

STU-PID: Steering Token Usage via PID Controller for Efficient Large Language Model Reasoning

ResearchDGX agent

arXiv:2506.18831v2 Announce Type: replace Abstract: Large Language Models employing extended chain-of-thought (CoT) reasoning often suffer from the overthinking phenomenon, generating excessive and re

TempusBench: An Evaluation Framework for Time-Series Forecasting

Model ReleasesDGX agent

arXiv:2604.11529v1 Announce Type: new Abstract: Foundation models have transformed natural language processing and computer vision, and a rapidly growing literature on time-series foundation models (T

Transformers Learn Latent Mixture Models In-Context via Mirror Descent

TutorialsDGX agent

arXiv:2604.10848v1 Announce Type: new Abstract: Sequence modelling requires determining which past tokens are causally relevant from the context and their importance: a process inherent to the attenti

Valence-Arousal Subspace in LLMs: Circular Emotion Geometry and Multi-Behavioral Control

Model ReleasesDGX agent

arXiv:2604.03147v2 Announce Type: replace-cross Abstract: We present a method to identify a valence-arousal (VA) subspace within large language model representations. From 211k emotion-labeled texts,

We may have a new SOTA open-source model: ERNIE-Image Comparisons

Local AiDGX agent

ERNIE-Image is an open-weight text-to-image generation model developed by Baidu, built on a single-stream Diffusion Transformer (DiT) paired with a lightweight Prompt Enhancer that expands brief user

What Do Vision-Language Models Encode for Personalized Image Aesthetics Assessment?

ApplicationsDGX agent

arXiv:2604.11374v1 Announce Type: cross Abstract: Personalized image aesthetics assessment (PIAA) is an important research problem with practical real-world applications. While methods based on vision

13 Apr 2026

A Mathematical Framework for Temporal Modeling and Counterfactual Policy Simulation of Student Dropout

SafetyDGX agent

arXiv:2604.08874v1 Announce Type: cross Abstract: This study proposes a temporal modeling framework with a counterfactual policy-simulation layer for student dropout in higher education, using LMS eng

AMO-ENE: Attention-based Multi-Omics Fusion Model for Outcome Prediction in Extra Nodal Extension and HPV-associated Oropharyngeal Cancer

ResearchDGX agent

arXiv:2604.09280v1 Announce Type: cross Abstract: Extranodal extension (ENE) is an emerging prognostic factor in human papillomavirus (HPV)-associated oropharyngeal cancer (OPC), although it is curren

An Adaptive Model Selection Framework for Demand Forecasting under Horizon-Induced Degradation to Support Business Strategy and Operations

SafetyDGX agent

arXiv:2602.13939v3 Announce Type: replace-cross Abstract: Business environments characterized by intermittent demand, high variability, and multi-step planning horizons require forecasting policies th

Bayesian Social Deduction with Graph-Informed Language Models

AgentsDGX agent

arXiv:2506.17788v2 Announce Type: replace Abstract: Social reasoning - inferring unobservable beliefs and intentions from partial observations of other agents - remains a challenging task for large la

Demystifying Mergeability: Interpretable Properties to Predict Model Merging Success

SafetyDGX agent

arXiv:2601.22285v4 Announce Type: replace Abstract: Model merging combines knowledge from separately fine-tuned models, yet success factors remain poorly understood. While recent work treats mergeabil

Graph Defense Diffusion Model

ApplicationsDGX agent

arXiv:2501.11568v2 Announce Type: replace Abstract: Graph Neural Networks (GNNs) are highly vulnerable to adversarial attacks, which can greatly degrade their performance. Existing graph purification

Listener-Rewarded Thinking in VLMs for Image Preferences

Model ReleasesDGX agent

arXiv:2506.22832v3 Announce Type: replace-cross Abstract: Training robust and generalizable reward models for human visual preferences is essential for aligning text-to-image and text-to-video generat

🦞👾 LM Studio is now an official @openclaw provider! Run: openclaw onboard --auth-choice lmstudio Use your local models with your OpenClaw …

Local AiDGX agent

🦞👾 LM Studio is now an official @openclaw provider! Run: openclaw onboard --auth-choice lmstudio Use your local models with your OpenClaw - it's private and free. Works on Mac, Windows, and Linux. Let

LMGenDrive: Bridging Multimodal Understanding and Generative World Modeling for End-to-End Driving

SafetyDGX agent

arXiv:2604.08719v1 Announce Type: cross Abstract: Recent years have seen remarkable progress in autonomous driving, yet generalization to long-tail and open-world scenarios remains a major bottleneck

M-IDoL: Information Decomposition for Modality-Specific and Diverse Representation Learning in Medical Foundation Model

TutorialsDGX agent

arXiv:2604.08936v1 Announce Type: new Abstract: Medical foundation models (MFMs) aim to learn universal representations from multimodal medical images that can generalize effectively to diverse downst

Model Space Reasoning as Search in Feedback Space for Planning Domain Generation

AgentsDGX agent

arXiv:2604.08712v1 Announce Type: new Abstract: The generation of planning domains from natural language descriptions remains an open problem even with the advent of large language models and reasonin

On the Limits of Layer Pruning for Generative Reasoning in Large Language Models

ResearchDGX agent

arXiv:2602.01997v2 Announce Type: replace-cross Abstract: Recent work has shown that layer pruning can effectively compress large language models (LLMs) while retaining strong performance on classific

Online3R: Online Learning for Consistent Sequential Reconstruction Based on Geometry Foundation Model

Local AiDGX agent

arXiv:2604.09480v1 Announce Type: new Abstract: We present Online3R, a new sequential reconstruction framework that is capable of adapting to new scenes through online learning, effectively resolving

PaceLLM: Brain-Inspired Large Language Models for Long-Context Understanding

ResearchDGX agent

arXiv:2506.17310v3 Announce Type: replace-cross Abstract: While Large Language Models (LLMs) demonstrate strong performance across domains, their long-context capabilities are limited by transient neu

Re-Mask and Redirect: Exploiting Denoising Irreversibility in Diffusion Language Models

SafetyDGX agent

arXiv:2604.08557v1 Announce Type: cross Abstract: Diffusion-based language models (dLLMs) generate text by iteratively denoising masked token sequences. We show that their safety alignment rests on a

Reasoning Models Will Sometimes Lie About Their Reasoning

ResearchDGX agent

arXiv:2601.07663v3 Announce Type: replace Abstract: Hint-based faithfulness evaluations have established that Large Reasoning Models (LRMs) may not say what they think: they do not always volunteer in

Run Unsloth Qwen3.5 ggufs or your finetuned version in Ollama

Local AiDGX agent

This Reddit post from r/ollama discusses how to run Unsloth's quantized GGUF versions of Qwen3.5 — Alibaba's model family including variants such as 35B-A3B, 27B, 122B-A10B, and smaller models like 0.

Seedance 2.0 is now live in ComfyUI for everyone. This state-of-the-art model brings us one step closer to production-quality video generati…

ApplicationsDGX agent

Seedance 2.0, a state-of-the-art video generation model, is now available in ComfyUI for all users. The integration marks a significant step toward production-quality video generation capabilities wit

WAND: Windowed Attention and Knowledge Distillation for Efficient Autoregressive Text-to-Speech Models

ResearchDGX agent

arXiv:2604.08558v1 Announce Type: cross Abstract: Recent decoder-only autoregressive text-to-speech (AR-TTS) models produce high-fidelity speech, but their memory and compute costs scale quadratically

12 Apr 2026

@hwchase17 is naming the lock-in play that model providers don't want builders to notice. Your agent's memory is the valuable part. Personal…

AgentsDGX agent

@hwchase17 is naming the lock-in play that model providers don't want builders to notice. Your agent's memory is the valuable part. Personalization, context, preferences that compound over time. If th

Model providers don’t lock you in with the API. They lock you in with your own data. Memory is the moat. If you don’t own your agent’s harne…

AgentsDGX agent

Model providers don’t lock you in with the API. They lock you in with your own data. Memory is the moat. If you don’t own your agent’s harness, you don’t own your agent’s memory. And switching means s

Suggestions on which model I should train an MC Escher Tessellation LoRA on?

Local AiDGX agent

This Reddit thread from r/StableDiffusion discusses community recommendations for choosing the best base model on which to train a LoRA (Low-Rank Adaptation) capturing M.C. Escher's distinctive tessel

11 Apr 2026

A 'Neural Computer' is built by adapting video generation architectures to train a World Model of an actual computer that can directly simul…

ResearchDGX agent

A 'Neural Computer' is built by adapting video generation architectures to train a World Model of an actual computer that can directly simulate a computer interface. Instead of interacting with a real

Black Forest Labs (@bfl_ml) Developer Advocate @stephenbtl says their first principle is building state-of-the-art models and shares why doi…

ToolsDGX agent

Black Forest Labs (@bfl_ml) Developer Advocate @stephenbtl says their first principle is building state-of-the-art models and shares why doing so from Freiburg, Germany is an advantage: 'We're from Fr

The inevitable need for an open model consortium

ResearchDGX agent

The article from Interconnects argues that the AI industry increasingly requires a formal consortium or collaborative organization dedicated to developing and maintaining open-weight language models,

10 Apr 2026

A GAN and LLM-Driven Data Augmentation Framework for Dynamic Linguistic Pattern Modeling in Chinese Sarcasm Detection

ResearchDGX agent

arXiv:2604.08381v1 Announce Type: new Abstract: Sarcasm is a rhetorical device that expresses criticism or emphasizes characteristics of certain individuals or situations through exaggeration, irony,

← Previous
1…144145146147148…1010
Next →