AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,002 results
14 Apr 2026

Revisiting Epistemic Markers in Confidence Estimation: Can Markers Accurately Reflect Large Language Models' Uncertainty?

SafetyDGX agent

arXiv:2505.24778v3 Announce Type: replace Abstract: As large language models (LLMs) are increasingly used in high-stakes domains, accurately assessing their confidence is crucial. Humans typically exp

RoboStereo: Dual-Tower 4D Embodied World Models for Unified Policy Optimization

SafetyDGX agent

arXiv:2603.12639v2 Announce Type: replace Abstract: Scalable Embodied AI faces fundamental constraints due to prohibitive costs and safety risks of real-world interaction. While Embodied World Models

SEPTQ: A Simple and Effective Post-Training Quantization Paradigm for Large Language Models

ResearchDGX agent

arXiv:2604.10091v1 Announce Type: new Abstract: Large language models (LLMs) have shown remarkable performance in various domains, but they are constrained by massive computational and storage costs.

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

SMFormer: Empowering Self-supervised Stereo Matching via Foundation Models and Data Augmentation

Model ReleasesDGX agent

arXiv:2604.10218v1 Announce Type: new Abstract: Recent self-supervised stereo matching methods have made significant progress. They typically rely on the photometric consistency assumption, which pres

StableTTA: Training-Free Test-Time Adaptation that Improves Model Accuracy on ImageNet1K to 96%

Model ReleasesDGX agent

arXiv:2604.04552v2 Announce Type: replace-cross Abstract: Ensemble methods are widely used to improve predictive performance, but their effectiveness often comes at the cost of increased memory usage

Towards Reasonable Concept Bottleneck Models

ResearchDGX agent

arXiv:2506.05014v2 Announce Type: replace-cross Abstract: We propose a novel, flexible, and efficient framework for designing Concept Bottleneck Models (CBMs) that enables practitioners to explicitly

Towards Situation-aware State Modeling for Air Traffic Flow Prediction

ApplicationsDGX agent

arXiv:2604.11198v1 Announce Type: new Abstract: Accurate air traffic prediction in the terminal airspace (TA) is pivotal for proactive air traffic management (ATM). However, existing data-driven appro

Vibe-driven model-based engineering

ResearchDGX agent

arXiv:2604.10645v1 Announce Type: cross Abstract: There is a pressing need for better development methods and tools to keep up with the growing demand and increasing complexity of new software systems

Waypoint-1.5 New open source world model trained on FPS games to run on local consumer GPUs at 60fps

Local AiDGX agent

Waypoint-1.5 New open source world model trained on FPS games to run on local consumer GPUs at 60fps

What do your logits know? (The answer may surprise you!)

TutorialsDGX agent

arXiv:2604.09885v1 Announce Type: new Abstract: Recent work has shown that probing model internals can reveal a wealth of information not apparent from the model generations. This poses the risk of un

⚡️ Zig 0.16 is out. And the new I/O model is a huge shift. • Swap implementations (threaded, evented, etc.) • Write code that looks blocking…

TutorialsDGX agent

⚡️ Zig 0.16 is out. And the new I/O model is a huge shift. • Swap implementations (threaded, evented, etc.) • Write code that looks blocking but runs async • Composable like allocators https://ziglang

13 Apr 2026

Did the $100 Plan Affect the GPT-5.4 Pro Model?

Model ReleasesDGX agent

This Reddit thread likely discusses community questions around OpenAI's new 100/month ChatGPT Pro tier and its implications for access to GPT-5.4 Pro. OpenAI introduced a 100/month Pro tier positioned

Dream to Fly: Model-Based Reinforcement Learning for Vision-Based Drone Flight

Model ReleasesDGX agent

arXiv:2501.14377v2 Announce Type: replace Abstract: Autonomous drone racing has risen as a challenging robotic benchmark for testing the limits of learning, perception, planning, and control. Expert h

DSVTLA: Deep Swin Vision Transformer-Based Transfer Learning Architecture for Multi-Type Cancer Histopathological Cancer Image Classification

Model ReleasesDGX agent

arXiv:2604.09468v1 Announce Type: cross Abstract: In this study, we proposed a deep Swin-Vision Transformer-based transfer learning architecture for robust multi-cancer histopathological image classif

Enhancing the Safety of Medical Vision-Language Models by Synthetic Demonstrations

SafetyDGX agent

arXiv:2506.09067v2 Announce Type: replace-cross Abstract: Generative medical vision-language models~(Med-VLMs) are primarily designed to generate complex textual information~(e.g., diagnostic reports)

GeRM: A Generative Rendering Model From Physically Realistic to Photorealistic

AgentsDGX agent

arXiv:2604.09304v1 Announce Type: new Abstract: For decades, Physically-Based Rendering (PBR) is the fundation of synthesizing photorealisitic images, and therefore sometimes roughly referred as Photo

Grammar as a Behavioral Biometric: Using Cognitively Motivated Grammar Models for Authorship Verification

ResearchDGX agent

arXiv:2403.08462v3 Announce Type: replace Abstract: Authorship Verification (AV) is a key area of research in digital text forensics, which addresses the fundamental question of whether two texts were

Hermes Agent Tip💡 Hermes supports dedicated auxiliary models for eight task types: 1. vision 2. web_extract 3. compression 4. session_searc…

AgentsDGX agent

Hermes Agent Tip💡 Hermes supports dedicated auxiliary models for eight task types: 1. vision 2. web_extract 3. compression 4. session_search 5. approval 6. skills_hub 7. mcp 8. flush_memories Each tas

InsEdit: Towards Instruction-based Visual Editing via Data-Efficient Video Diffusion Models Adaptation

ResearchDGX agent

arXiv:2604.08646v1 Announce Type: new Abstract: Instruction-based video editing is a natural way to control video content with text, but adapting a video generation model into an editor usually appear

Koopman Operator Framework for Modeling and Control of Off-Road Vehicle on Deformable Terrain

AgentsDGX agent

arXiv:2603.28965v2 Announce Type: replace-cross Abstract: This work presents a hybrid physics-informed and data-driven modeling framework for predictive control of autonomous off-road vehicles operati

Large Reasoning Models Learn Better Alignment from Flawed Thinking

SafetyDGX agent

arXiv:2510.00938v2 Announce Type: replace Abstract: Large reasoning models (LRMs) 'think' by generating structured chain-of-thought (CoT) before producing a final answer, yet they still lack the abili

Parameterized Complexity Of Representing Models Of MSO Formulas

Model ReleasesDGX agent

arXiv:2604.08707v1 Announce Type: new Abstract: Monadic second order logic (MSO2) plays an important role in parameterized complexity due to the Courcelle's theorem. This theorem states that the probl

PRAGMA: Revolut Foundation Model

ResearchDGX agent

arXiv:2604.08649v1 Announce Type: cross Abstract: Modern financial systems generate vast quantities of transactional and event-level data that encode rich economic signals. This paper presents PRAGMA,

QoS-QoE Translation with Large Language Model

Model ReleasesDGX agent

arXiv:2604.08703v1 Announce Type: cross Abstract: QoS-QoE translation is a fundamental problem in multimedia systems because it characterizes how measurable system and network conditions affect user-p

Robust Reasoning Benchmark

Model ReleasesDGX agent

arXiv:2604.08571v1 Announce Type: cross Abstract: While Large Language Models (LLMs) achieve high performance on standard mathematical benchmarks, their underlying reasoning processes remain highly ov

Tango: Taming Visual Signals for Efficient Video Large Language Models

Local AiDGX agent

arXiv:2604.09547v1 Announce Type: new Abstract: Token pruning has emerged as a mainstream approach for developing efficient Video Large Language Models (Video LLMs). This work revisits and advances th

Token Reduction via Local and Global Contexts Optimization for Efficient Video Large Language Models

ResearchDGX agent

arXiv:2603.01400v2 Announce Type: replace Abstract: Video Large Language Models (VLLMs) demonstrate strong video understanding but suffer from inefficiency due to redundant visual tokens. Existing pru

We wrote a build-from-scratch Python book on the 98% of production AI systems that isn't the model call [P]

ApplicationsDGX agent

This Reddit post from r/MachineLearning announces a build-from-scratch Python book focused on the engineering infrastructure surrounding AI systems — the components beyond the model call itself, such

What Matters in Virtual Try-Off? Dual-UNet Diffusion Model For Garment Reconstruction

TutorialsDGX agent

arXiv:2604.08716v1 Announce Type: new Abstract: Virtual Try-On (VTON) has seen rapid advancements, providing a strong foundation for generative fashion tasks. However, the inverse problem, Virtual Try

12 Apr 2026

gemma 4 going insane

Model ReleasesDGX agent

This r/ollama thread discusses users experiencing erratic and broken behavior when running Google's Gemma 4 models locally via Ollama, including issues such as strange and unrelated responses, as well

model page: https://ollama.com/library/minimax-m2.7

Local AiDGX agent

MiniMax-M1 is a large-scale open-weight reasoning model developed by MiniMax, featuring a hybrid Mixture-of-Experts (MoE) architecture with 456 billion total parameters (activating 45.9 billion per to

the most important abstraction in AI agents isnt the model — its the harness it orchestrates tools, memory, prompts. this is where all the a…

AgentsDGX agent

the most important abstraction in AI agents isnt the model — its the harness it orchestrates tools, memory, prompts. this is where all the alpha is deepagents is our take: built-in tools, memory, smar

11 Apr 2026

AI models are terrible at betting on soccer—especially xAI Grok

IndustryDGX agent

A report called 'KellyBench,' released by AI start-up General Reasoning, tested eight leading AI models in a virtual re-creation of the 2023–24 Premier League season, providing them with detailed hist

Does UI Preset = Base model??

Local AiDGX agent

This Reddit thread from r/StableDiffusion addresses a common point of confusion among users of Stable Diffusion UIs (such as Stable Diffusion WebUI Forge) regarding whether selecting a 'UI Preset' is

10 Apr 2026

A systematic framework for generating novel experimental hypotheses from language models

SafetyDGX agent

arXiv:2408.05086v3 Announce Type: replace Abstract: Neural language models (LMs) have been shown to capture complex linguistic patterns, yet their utility in understanding human language and more broa

ABMAMBA: Multimodal Large Language Model with Aligned Hierarchical Bidirectional Scan for Efficient Video Captioning

ResearchDGX agent

arXiv:2604.08050v1 Announce Type: new Abstract: In this study, we focus on video captioning by fully open multimodal large language models (MLLMs). The comprehension of visual sequences is challenging

AFL: A Single-Round Analytic Approach for Federated Learning with Pre-trained Models

ResearchDGX agent

arXiv:2405.16240v3 Announce Type: replace Abstract: In this paper, we introduce analytic federated learning (AFL), a new training paradigm that brings analytical (i.e., closed-form) solutions to the f

Apple's head of cloud says Open Source models will address 90% of the use case

IndustryDGX agent

I was unable to retrieve the specific Reddit thread or locate reliable sourced details about Apple's head of cloud making a statement that open source models will address 90% of use cases. The sear...

Attribution-Driven Explainable Intrusion Detection with Encoder-Based Large Language Models

TutorialsDGX agent

arXiv:2604.06266v1 Announce Type: cross Abstract: Software-Defined Networking (SDN) improves network flexibility but also increases the need for reliable and interpretable intrusion detection. Large L

Beyond the Mean: Modelling Annotation Distributions in Continuous Affect Prediction

ResearchDGX agent

arXiv:2604.07198v1 Announce Type: new Abstract: Emotion annotation is inherently subjective and cognitively demanding, producing signals that reflect diverse perceptions across annotators rather than

DISSECT: Diagnosing Where Vision Ends and Language Priors Begin in Scientific VLMs

Model ReleasesDGX agent

arXiv:2604.06250v1 Announce Type: cross Abstract: When asked to describe a molecular diagram, a Vision-Language Model correctly identifies ``a benzene ring with an -OH group.'' When asked to reason ab

'Don't Do That!': Guiding Embodied Systems through Large Language Model-based Constraint Generation

ResearchDGX agent

arXiv:2506.04500v3 Announce Type: replace-cross Abstract: Recent advancements in large language models (LLMs) have spurred interest in robotic navigation that incorporates complex spatial, mathematica

Energy-Regularized Spatial Masking: A Novel Approach to Enhancing Robustness and Interpretability in Vision Models

ResearchDGX agent

arXiv:2604.06893v1 Announce Type: cross Abstract: Deep convolutional neural networks achieve remarkable performance by exhaustively processing dense spatial feature maps, yet this brute-force strategy

Flow Motion Policy: Manipulator Motion Planning with Flow Matching Models

Model ReleasesDGX agent

arXiv:2604.07084v1 Announce Type: cross Abstract: Open-loop end-to-end neural motion planners have recently been proposed to improve motion planning for robotic manipulators. These methods enable plan

Hallucination as output-boundary misclassification: a composite abstention architecture for language models

ResearchDGX agent

arXiv:2604.06195v1 Announce Type: cross Abstract: Large language models often produce unsupported claims. We frame this as a misclassification error at the output boundary, where internally generated

Hierarchical Feature Learning for Medical Point Clouds via State Space Model

ResearchDGX agent

arXiv:2504.13015v3 Announce Type: replace Abstract: Deep learning-based point cloud modeling has been widely investigated as an indispensable component of general shape analysis. Recently, transformer

Incentive-Aware Multi-Fidelity Optimization for Generative Advertising in Large Language Models

ResearchDGX agent

arXiv:2604.06263v1 Announce Type: cross Abstract: Generative advertising in large language model (LLM) responses requires optimizing sponsorship configurations under two strict constraints: the strate

Learning is Forgetting: LLM Training As Lossy Compression

TutorialsDGX agent

arXiv:2604.07569v1 Announce Type: cross Abstract: Despite the increasing prevalence of large language models (LLMs), we still have a limited understanding of how their representational spaces are stru

LINE: LLM-based Iterative Neuron Explanations for Vision Models

SafetyDGX agent

arXiv:2604.08039v1 Announce Type: new Abstract: Interpreting the concepts encoded by individual neurons in deep neural networks is a crucial step towards understanding their complex decision-making pr

LongWriter-Zero: Mastering Ultra-Long Text Generation via Reinforcement Learning

Model ReleasesDGX agent

arXiv:2506.18841v3 Announce Type: replace-cross Abstract: Ultra-long generation by large language models (LLMs) is a widely demanded scenario, yet it remains a significant challenge due to their maxim

ODE-free Neural Flow Matching for One-Step Generative Modeling

TutorialsDGX agent

arXiv:2604.06413v1 Announce Type: new Abstract: Diffusion and flow matching models generate samples by learning time-dependent vector fields whose integration transports noise to data, requiring tens

One Life to Learn: Inferring Symbolic World Models for Stochastic Environments from Unguided Exploration

AgentsDGX agent

arXiv:2510.12088v2 Announce Type: replace Abstract: Symbolic world modeling requires inferring and representing an environment's transitional dynamics as an executable program. Prior work has focused

Rethinking Entropy Allocation in LLM-based ASR: Understanding the Dynamics between Speech Encoders and LLMs

Model ReleasesDGX agent

arXiv:2604.08003v1 Announce Type: cross Abstract: Integrating large language models (LLMs) into automatic speech recognition (ASR) has become a dominant paradigm. Although recent LLM-based ASR models

STQuant: Spatio-Temporal Adaptive Framework for Optimizer Quantization in Large Multimodal Model Training

ResearchDGX agent

arXiv:2604.06836v2 Announce Type: new Abstract: Quantization is an effective way to reduce the memory cost of large-scale model training. However, most existing methods adopt fixed-precision policies,

The quality of talks at @aiDotEngineer is insane, being able to learn about diffusion models and flow mapping from @GoogleDeepMind’s @sediel…

TutorialsDGX agent

The AI Engineer Summit (@aiDotEngineer) is a highly regarded technical conference featuring speakers from leading AI organizations including Google DeepMind, Anthropic, and OpenAI, known for its de...

Visual prompting reimagined: The power of the Activation Prompts

Model ReleasesDGX agent

arXiv:2604.06440v1 Announce Type: cross Abstract: Visual prompting (VP) has emerged as a popular method to repurpose pretrained vision models for adaptation to downstream tasks. Unlike conventional mo

9 Apr 2026

a useful mental model on how teams can think about good data design to improve their models/agents: Evals ~= Training Data ~= Environments -…

AgentsDGX agent

a useful mental model on how teams can think about good data design to improve their models/agents: Evals ~= Training Data ~= Environments - in Classical Deep Learning, we learn from each training exa

Another banger paper from Microsoft. Why it's a big deal: It teaches reasoning models to compress their own chain-of-thought mid-generation.…

AgentsDGX agent

Another banger paper from Microsoft. Why it's a big deal: It teaches reasoning models to compress their own chain-of-thought mid-generation. The most interesting finding isn't the 2-3x memory savings

This 10-min read from @Vtrivedy10 changes how you build AI agents. Most people are stuck in the same loop; switching models when agents brea…

AgentsDGX agent

This 10-min read from @Vtrivedy10 changes how you build AI agents. Most people are stuck in the same loop; switching models when agents break. The reframe: evals are the training data for your harness

8 Apr 2026

Anthropic had the most powerful cyber-security model in the history of this world and their internal code based still leaked? We should assu…

IndustryDGX agent

Anthropic had the most powerful cyber-security model in the history of this world and their internal code based still leaked? We should assume everyone can be compromised, and build systems that keep

← Previous
1…168169170171172…1017
Next →