AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,575 results
15 Apr 2026

Human-Centric Topic Modeling with Goal-Prompted Contrastive Learning and Optimal Transport

SafetyDGX agent

arXiv:2604.12663v1 Announce Type: new Abstract: Existing topic modeling methods, from LDA to recent neural and LLM-based approaches, which focus mainly on statistical coherence, often produce redundan

I spent some time trying to distill all the complex factors impacting open models -- economics, capabilities, distribution, policy, etc. -- …

SafetyDGX agent

I spent some time trying to distill all the complex factors impacting open models -- economics, capabilities, distribution, policy, etc. -- into a clear list of beliefs. Here they are in full. 1. It’s

KG-Reasoner: A Reinforced Model for End-to-End Multi-Hop Knowledge Graph Reasoning

ResearchDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.12487v1 Announce Type: cross Abstract: Large Language Models (LLMs) exhibit strong abilities in natural language understanding and generation, yet they struggle with knowledge-intensive rea

Knowledge Is Not Static: Order-Aware Hypergraph RAG for Language Models

ApplicationsDGX agent

arXiv:2604.12185v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) enhances large language models by grounding outputs in retrieved knowledge. However, existing RAG methods including

Latent-Condensed Transformer for Efficient Long Context Modeling

ResearchDGX agent

arXiv:2604.12452v1 Announce Type: new Abstract: Large language models (LLMs) face significant challenges in processing long contexts due to the linear growth of the key-value (KV) cache and quadratic

Models Know Their Shortcuts: Deployment-Time Shortcut Mitigation

SafetyDGX agent

arXiv:2604.12277v1 Announce Type: new Abstract: Pretrained language models often rely on superficial features that appear predictive during training yet fail to generalize at test time, a phenomenon k

More people should work on harnesses for open and local models!

IndustryDGX agent

Hugging Face co-founder and CEO Clément Delangue advocates for more developers and researchers to focus on building evaluation harnesses and testing frameworks specifically designed for open-source an

Parcae: Scaling Laws For Stable Looped Language Models

Model ReleasesDGX agent

arXiv:2604.12946v1 Announce Type: new Abstract: Traditional fixed-depth architectures scale quality by increasing training FLOPs, typically through increased parameterization, at the expense of a high

Point Prompting: Counterfactual Tracking with Video Diffusion Models

ResearchDGX agent

arXiv:2510.11715v2 Announce Type: replace Abstract: Trackers and video generators solve closely related problems: the former analyze motion, while the latter synthesize it. We show that this connectio

ReasonXL: Shifting LLM Reasoning Language Without Sacrificing Performance

Model ReleasesDGX agent

arXiv:2604.12378v1 Announce Type: new Abstract: Despite advances in multilingual capabilities, most large language models (LLMs) remain English-centric in their training and, crucially, in their produ

ResBM: Residual Bottleneck Models for Low-Bandwidth Pipeline Parallelism

ResearchDGX agent

arXiv:2604.11947v1 Announce Type: cross Abstract: Unlocking large-scale low-bandwidth decentralized training has the potential to utilize otherwise untapped compute resources. In centralized settings,

Siamese Foundation Models for Crystal Structure Prediction

TutorialsDGX agent

arXiv:2503.10471v2 Announce Type: replace-cross Abstract: Predicting crystal structures from chemical compositions is a fundamental challenge in materials discovery, complicated by complex 3D geometri

SparseWorld-TC: Trajectory-Conditioned Sparse Occupancy World Model

Model ReleasesDGX agent

arXiv:2511.22039v3 Announce Type: replace Abstract: This paper introduces a novel architecture for trajectory-conditioned forecasting of future 3D scene occupancy. In contrast to methods that rely on

The A-R Behavioral Space: Execution-Level Profiling of Tool-Using Language Model Agents in Organizational Deployment

SafetyDGX agent

arXiv:2604.12116v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed as tool-augmented agents capable of executing system-level operations. While existing benchmarks

VULCAN: Vision-Language-Model Enhanced Multi-Agent Cooperative Navigation for Indoor Fire-Disaster Response

Model ReleasesDGX agent

arXiv:2604.12831v1 Announce Type: new Abstract: Indoor fire disasters pose severe challenges to autonomous search and rescue due to dense smoke, high temperatures, and dynamically evolving indoor envi

WikiSeeker: Rethinking the Role of Vision-Language Models in Knowledge-Based Visual Question Answering

Model ReleasesDGX agent

arXiv:2604.05818v2 Announce Type: replace-cross Abstract: Multi-modal Retrieval-Augmented Generation (RAG) has emerged as a highly effective paradigm for Knowledge-Based Visual Question Answering (KB-

14 Apr 2026

A robust and adaptive MPC formulation for Gaussian process models

TutorialsDGX agent

arXiv:2507.02098v2 Announce Type: replace-cross Abstract: In this paper, we present a robust and adaptive model predictive control (MPC) framework for uncertain nonlinear systems affected by bounded d

ASPIRin: Action Space Projection for Interactivity-Optimized Reinforcement Learning in Full-Duplex Speech Language Models

SafetyDGX agent

arXiv:2604.10065v1 Announce Type: cross Abstract: End-to-end full-duplex Speech Language Models (SLMs) require precise turn-taking for natural interaction. However, optimizing temporal dynamics via st

BoxTuning: Directly Injecting the Object Box for Multimodal Model Fine-Tuning

ResearchDGX agent

arXiv:2604.11136v1 Announce Type: cross Abstract: Object-level spatial-temporal understanding is essential for video question answering, yet existing multimodal large language models (MLLMs) encode fr

ChemPro: A Progressive Chemistry Benchmark for Large Language Models

Model ReleasesDGX agent

arXiv:2602.03108v3 Announce Type: replace Abstract: We introduce ChemPro, a progressive benchmark with 4100 natural language question-answer pairs in Chemistry, across 4 coherent sections of difficult

CLSGen: A Dual-Head Fine-Tuning Framework for Joint Probabilistic Classification and Verbalized Explanation

Model ReleasesDGX agent

arXiv:2604.11801v1 Announce Type: new Abstract: With the recent progress of Large Language Models (LLMs), there is a growing interest in applying these models to solve complex and challenging problems

CoSToM:Causal-oriented Steering for Intrinsic Theory-of-Mind Alignment in Large Language Models

SafetyDGX agent

arXiv:2604.10031v1 Announce Type: cross Abstract: Theory of Mind (ToM), the ability to attribute mental states to others, is a hallmark of social intelligence. While large language models (LLMs) demon

DeCoVec: Building Decoding Space based Task Vector for Large Language Models via In-Context Learning

ResearchDGX agent

arXiv:2604.11129v1 Announce Type: new Abstract: Task vectors, representing directions in model or activation spaces that encode task-specific behaviors, have emerged as a promising tool for steering l

dTRPO: Trajectory Reduction in Policy Optimization of Diffusion Large Language Models

SafetyDGX agent

arXiv:2603.18806v2 Announce Type: replace Abstract: Diffusion Large Language Models (dLLMs) introduce a new paradigm for language generation, which in turn presents new challenges for aligning them wi

FineEdit: Fine-Grained Image Edit with Bounding Box Guidance

Model ReleasesDGX agent

arXiv:2604.10954v1 Announce Type: new Abstract: Diffusion-based image editing models have achieved significant progress in real world applications. However, conventional models typically rely on natur

Fixed: IPEX-LLM + modern Ollama models (qwen3, gemma4) on Intel Arc 140V Lunar Lake Windows 11 — undocumented solution

Local AiDGX agent

This Reddit post from r/ollama documents a community-discovered, undocumented workaround for getting IPEX-LLM to successfully run modern Ollama models — specifically Qwen3 and Gemma4 — on systems powe

For model details: https://blog.comfy.org/p/comfyui-now-supports-sonilo-via-partner

Local AiDGX agent

ComfyUI has announced support for Sonilo, a new AI model integrated through a partner collaboration, with full technical details available on the ComfyUI blog. The integration allows users to run Soni

GlobalCY I: A JAX Framework for Globally Defined and Symmetry-Aware Neural Kahler Potentials

Model ReleasesDGX agent

arXiv:2604.11404v1 Announce Type: cross Abstract: We present GlobalCY, a JAX-based framework for globally defined and symmetry-aware neural Kahler-potential models on projective hypersurface Calabi--Y

Honest question - What model are Iran using for those excellent Lego Videos?

Local AiDGX agent

This Reddit thread from r/StableDiffusion discusses the AI tools behind the viral Lego-style propaganda videos produced by Explosive Media, known in Persian as Akhbar Enfejari — a group whose AI-gener

HuiYanEarth-SAR: A Foundation Model for High-Fidelity and Low-Cost Global Remote Sensing Imagery Generation

ResearchDGX agent

arXiv:2604.11444v1 Announce Type: new Abstract: Synthetic Aperture Radar (SAR) imagery generation is essential for deepening the study of scattering mechanisms, establishing trustworthy electromagneti

If an LLM Were a Character, Would It Know Its Own Story? Evaluating Lifelong Learning in LLMs

Model ReleasesDGX agent

arXiv:2503.23514v2 Announce Type: replace-cross Abstract: Large language models (LLMs) can carry out human-like dialogue, but unlike humans, they are stateless due to the superposition property. Howev

Intelligent Approval of Access Control Flow in Office Automation Systems via Relational Modeling

ApplicationsDGX agent

arXiv:2604.11040v1 Announce Type: new Abstract: Office automation (OA) systems play a crucial role in enterprise operations and management, with access control flow approval (ACFA) being a key compone

ITIScore: An Image-to-Text-to-Image Rating Framework for the Image Captioning Ability of MLLMs

Model ReleasesDGX agent

arXiv:2604.03765v2 Announce Type: replace Abstract: Recent advances in multimodal large language models (MLLMs) have greatly improved image understanding and captioning capabilities. However, existing

Learning Racket-Ball Bounce Dynamics Across Diverse Rubbers for Robotic Table Tennis

Model ReleasesDGX agent

arXiv:2604.11349v1 Announce Type: new Abstract: Accurate dynamic models for racket-ball bounces are essential for reliable control in robotic table tennis. Existing models typically assume simple line

New Hybrid Fine-Tuning Paradigm for LLMs: Algorithm Design and Convergence Analysis Framework

Model ReleasesDGX agent

arXiv:2604.09940v1 Announce Type: new Abstract: Fine-tuning Large Language Models (LLMs) typically involves either full fine-tuning, which updates all model parameters, or Parameter-Efficient Fine-Tun

One Scale at a Time: Scale-Autoregressive Modeling for Fluid Flow Distributions

ApplicationsDGX agent

arXiv:2604.11403v1 Announce Type: cross Abstract: Analyzing unsteady fluid flows often requires access to the full distribution of possible temporal states, yet conventional PDE solvers are computatio

Revisiting Epistemic Markers in Confidence Estimation: Can Markers Accurately Reflect Large Language Models' Uncertainty?

SafetyDGX agent

arXiv:2505.24778v3 Announce Type: replace Abstract: As large language models (LLMs) are increasingly used in high-stakes domains, accurately assessing their confidence is crucial. Humans typically exp

RoboStereo: Dual-Tower 4D Embodied World Models for Unified Policy Optimization

SafetyDGX agent

arXiv:2603.12639v2 Announce Type: replace Abstract: Scalable Embodied AI faces fundamental constraints due to prohibitive costs and safety risks of real-world interaction. While Embodied World Models

SEPTQ: A Simple and Effective Post-Training Quantization Paradigm for Large Language Models

ResearchDGX agent

arXiv:2604.10091v1 Announce Type: new Abstract: Large language models (LLMs) have shown remarkable performance in various domains, but they are constrained by massive computational and storage costs.

SMFormer: Empowering Self-supervised Stereo Matching via Foundation Models and Data Augmentation

Model ReleasesDGX agent

arXiv:2604.10218v1 Announce Type: new Abstract: Recent self-supervised stereo matching methods have made significant progress. They typically rely on the photometric consistency assumption, which pres

StableTTA: Training-Free Test-Time Adaptation that Improves Model Accuracy on ImageNet1K to 96%

Model ReleasesDGX agent

arXiv:2604.04552v2 Announce Type: replace-cross Abstract: Ensemble methods are widely used to improve predictive performance, but their effectiveness often comes at the cost of increased memory usage

Towards Reasonable Concept Bottleneck Models

ResearchDGX agent

arXiv:2506.05014v2 Announce Type: replace-cross Abstract: We propose a novel, flexible, and efficient framework for designing Concept Bottleneck Models (CBMs) that enables practitioners to explicitly

Towards Situation-aware State Modeling for Air Traffic Flow Prediction

ApplicationsDGX agent

arXiv:2604.11198v1 Announce Type: new Abstract: Accurate air traffic prediction in the terminal airspace (TA) is pivotal for proactive air traffic management (ATM). However, existing data-driven appro

Vibe-driven model-based engineering

ResearchDGX agent

arXiv:2604.10645v1 Announce Type: cross Abstract: There is a pressing need for better development methods and tools to keep up with the growing demand and increasing complexity of new software systems

Waypoint-1.5 New open source world model trained on FPS games to run on local consumer GPUs at 60fps

Local AiDGX agent

Waypoint-1.5 New open source world model trained on FPS games to run on local consumer GPUs at 60fps

What do your logits know? (The answer may surprise you!)

TutorialsDGX agent

arXiv:2604.09885v1 Announce Type: new Abstract: Recent work has shown that probing model internals can reveal a wealth of information not apparent from the model generations. This poses the risk of un

⚡️ Zig 0.16 is out. And the new I/O model is a huge shift. • Swap implementations (threaded, evented, etc.) • Write code that looks blocking…

TutorialsDGX agent

⚡️ Zig 0.16 is out. And the new I/O model is a huge shift. • Swap implementations (threaded, evented, etc.) • Write code that looks blocking but runs async • Composable like allocators https://ziglang

13 Apr 2026

Did the $100 Plan Affect the GPT-5.4 Pro Model?

Model ReleasesDGX agent

This Reddit thread likely discusses community questions around OpenAI's new 100/month ChatGPT Pro tier and its implications for access to GPT-5.4 Pro. OpenAI introduced a 100/month Pro tier positioned

Dream to Fly: Model-Based Reinforcement Learning for Vision-Based Drone Flight

Model ReleasesDGX agent

arXiv:2501.14377v2 Announce Type: replace Abstract: Autonomous drone racing has risen as a challenging robotic benchmark for testing the limits of learning, perception, planning, and control. Expert h

DSVTLA: Deep Swin Vision Transformer-Based Transfer Learning Architecture for Multi-Type Cancer Histopathological Cancer Image Classification

Model ReleasesDGX agent

arXiv:2604.09468v1 Announce Type: cross Abstract: In this study, we proposed a deep Swin-Vision Transformer-based transfer learning architecture for robust multi-cancer histopathological image classif

Enhancing the Safety of Medical Vision-Language Models by Synthetic Demonstrations

SafetyDGX agent

arXiv:2506.09067v2 Announce Type: replace-cross Abstract: Generative medical vision-language models~(Med-VLMs) are primarily designed to generate complex textual information~(e.g., diagnostic reports)

GeRM: A Generative Rendering Model From Physically Realistic to Photorealistic

AgentsDGX agent

arXiv:2604.09304v1 Announce Type: new Abstract: For decades, Physically-Based Rendering (PBR) is the fundation of synthesizing photorealisitic images, and therefore sometimes roughly referred as Photo

Grammar as a Behavioral Biometric: Using Cognitively Motivated Grammar Models for Authorship Verification

ResearchDGX agent

arXiv:2403.08462v3 Announce Type: replace Abstract: Authorship Verification (AV) is a key area of research in digital text forensics, which addresses the fundamental question of whether two texts were

Hermes Agent Tip💡 Hermes supports dedicated auxiliary models for eight task types: 1. vision 2. web_extract 3. compression 4. session_searc…

AgentsDGX agent

Hermes Agent Tip💡 Hermes supports dedicated auxiliary models for eight task types: 1. vision 2. web_extract 3. compression 4. session_search 5. approval 6. skills_hub 7. mcp 8. flush_memories Each tas

InsEdit: Towards Instruction-based Visual Editing via Data-Efficient Video Diffusion Models Adaptation

ResearchDGX agent

arXiv:2604.08646v1 Announce Type: new Abstract: Instruction-based video editing is a natural way to control video content with text, but adapting a video generation model into an editor usually appear

Koopman Operator Framework for Modeling and Control of Off-Road Vehicle on Deformable Terrain

AgentsDGX agent

arXiv:2603.28965v2 Announce Type: replace-cross Abstract: This work presents a hybrid physics-informed and data-driven modeling framework for predictive control of autonomous off-road vehicles operati

Large Reasoning Models Learn Better Alignment from Flawed Thinking

SafetyDGX agent

arXiv:2510.00938v2 Announce Type: replace Abstract: Large reasoning models (LRMs) 'think' by generating structured chain-of-thought (CoT) before producing a final answer, yet they still lack the abili

Parameterized Complexity Of Representing Models Of MSO Formulas

Model ReleasesDGX agent

arXiv:2604.08707v1 Announce Type: new Abstract: Monadic second order logic (MSO2) plays an important role in parameterized complexity due to the Courcelle's theorem. This theorem states that the probl

PRAGMA: Revolut Foundation Model

ResearchDGX agent

arXiv:2604.08649v1 Announce Type: cross Abstract: Modern financial systems generate vast quantities of transactional and event-level data that encode rich economic signals. This paper presents PRAGMA,

QoS-QoE Translation with Large Language Model

Model ReleasesDGX agent

arXiv:2604.08703v1 Announce Type: cross Abstract: QoS-QoE translation is a fundamental problem in multimedia systems because it characterizes how measurable system and network conditions affect user-p

← Previous
1…166167168169170…1010
Next →