AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

85,202Total entries
1Added by human
85,201Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,047 results
1 May 2026

Echo-{alpha}: Large Agentic Multimodal Reasoning Model for Ultrasound Interpretation

Local AiDGX agent

arXiv:2604.28011v1 Announce Type: new Abstract: Ultrasound interpretation requires both precise lesion localization and holistic clinical reasoning, yet existing methods typically excel at only one of

Exploring the Adoption Intention in Using AI-Enabled Educational Tools Among Preservice Teachers in the Philippines: A Partial-Least Square Modeling

ResearchDGX agent

arXiv:2604.27346v1 Announce Type: cross Abstract: This study examines the factors influencing pre-service teachers' behavioral intention to use AI-enabled educational tools during their practicum, usi

Flying by Inference: Active Inference World Models for Adaptive UAV Swarms

AgentsDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.27935v1 Announce Type: new Abstract: This paper presents an expert-guided active-inference-inspired framework for adaptive UAV swarm trajectory planning. The proposed method converts multi-

I have been testing DeepSeek-V4-Pro with the Pi coding agent. I am mindblown by how well it works out of the box. A few notes: I spent a few…

Model ReleasesDGX agent

I have been testing DeepSeek-V4-Pro with the Pi coding agent. I am mindblown by how well it works out of the box. A few notes: I spent a few hours building an LLM wiki with an agent powered entirely b

ImagineNav++: Prompting Vision-Language Models as Embodied Navigator through Scene Imagination

AgentsDGX agent

arXiv:2512.17435v3 Announce Type: replace Abstract: Visual navigation is a fundamental capability for autonomous home-assistance robots, enabling long-horizon tasks such as object search. While recent

KellyBench: A Benchmark for Long-Horizon Sequential Decision Making

Model ReleasesDGX agent

arXiv:2604.27865v1 Announce Type: new Abstract: Language models are saturating benchmarks for procedural tasks with narrow objectives. But they are increasingly being deployed in long-horizon, non-sta

Qualitative Evaluation of Language Model Rescoring in Automatic Speech Recognition

ResearchDGX agent

arXiv:2604.27533v1 Announce Type: new Abstract: Evaluating automatic speech recognition (ASR) systems is a classical but difficult and still open problem, which often boils down to focusing only on th

RayFormer: Modeling Inter- and Intra-Ray Similarity for NeRF-Based Video Snapshot Compressive Imaging

ApplicationsDGX agent

arXiv:2604.27702v1 Announce Type: new Abstract: Video snapshot compressive imaging (SCI) enables the reconstruction of dynamic scenes from a single snapshot measurement. Recently, NeRF-based methods h

Revealing the Impact of Visual Text Style on Attribute-based Descriptions Produced by Large Visual Language Models

ResearchDGX agent

arXiv:2604.27553v1 Announce Type: new Abstract: When the visual style of text is considered, a wide variety can be observed in font, color, and size. However, when a word is read, its meaning is indep

Standard Intelligence raises $75M to develop efficient computer use models

IndustryDGX agent

Standard Intelligence Inc., a six-person artificial intelligence startup, today announced that it has raised 75 million in funding. Sequoia and Spark Capital led the round. They were joined by multipl

30 Apr 2026

Exploring the Potential of Probabilistic Transformer for Time Series Modeling: A Report on the ST-PT Framework

ResearchDGX agent

arXiv:2604.26762v1 Announce Type: cross Abstract: The Probabilistic Transformer (PT) establishes that the Transformer's self-attention plus its feed-forward block is mathematically equivalent to Mean-

ext{PKS}^4:Parallel Kinematic Selective State Space Scanners for Efficient Video Understanding

Model ReleasesDGX agent

arXiv:2604.26461v1 Announce Type: new Abstract: Temporal modeling remains a fundamental challenge in video understanding, particularly as sequence lengths scale. Traditional video models relying on de

Inferix: A Block-Diffusion based Next-Generation Inference Engine for World Simulation

Model ReleasesDGX agent

arXiv:2511.20714v2 Announce Type: replace-cross Abstract: World models serve as core simulators for fields such as agentic AI, embodied AI, and gaming, capable of generating long, physically realistic

Integrating Weather Foundation Model and Satellite to Enable Fine-Grained Solar Irradiance Forecasting

ResearchDGX agent

arXiv:2603.14845v3 Announce Type: replace-cross Abstract: Accurate day-ahead solar irradiance forecasting is essential for integrating solar energy into the power grid. However, it remains challenging

PATCH: Learnable Tile-level Hybrid Sparsity for LLMs

Model ReleasesDGX agent

arXiv:2509.23410v4 Announce Type: replace-cross Abstract: Large language models (LLMs) deliver impressive performance but incur prohibitive memory and compute costs at deployment. Model pruning is an

Preserving Disagreement: Architectural Heterogeneity and Coherence Validation in Multi-Agent Policy Simulation

Model ReleasesDGX agent

arXiv:2604.26561v1 Announce Type: cross Abstract: Multi-agent deliberation systems using large language models (LLMs) are increasingly proposed for policy simulation, yet they suffer from artificial c

resharing this note, find it helpful given all the great open evals work + teams building vertical agents Evals are a proxy for the behavior…

Model ReleasesDGX agent

resharing this note, find it helpful given all the great open evals work + teams building vertical agents Evals are a proxy for the behavior we want our agent to exhibit in production Model+Harness pu

Training-Free Loosely Speculative Decoding: Accepting Semantically Correct Drafts Beyond Exact Match

Model ReleasesDGX agent

arXiv:2511.22972v3 Announce Type: replace Abstract: Large language models (LLMs) achieve strong performance across diverse tasks but suffer from high inference latency due to their autoregressive gene

29 Apr 2026

AdaTooler-V: Adaptive Tool-Use for Images and Videos

Model ReleasesDGX agent

arXiv:2512.16918v3 Announce Type: replace Abstract: Recent advances have shown that multimodal large language models (MLLMs) benefit from multimodal interleaved chain-of-thought (CoT) with vision tool

DRAGON: A Benchmark for Evidence-Grounded Visual Reasoning over Diagrams

Model ReleasesDGX agent

arXiv:2604.25231v1 Announce Type: cross Abstract: Diagram question answering (DQA) requires models to interpret structured visual representations such as charts, maps, infographics, circuit schematics

Incompressible Knowledge Probes: Estimating Black-Box LLM Parameter Counts via Factual Capacity

Model ReleasesDGX agent

arXiv:2604.24827v1 Announce Type: new Abstract: Closed-source frontier labs do not disclose parameter counts, and the standard alternative -- inference economics -- carries 2imes+ uncertainty from har

Large Language Models Are Effective Human Annotation Assistants, But Not Good Independent Annotators

ResearchDGX agent

arXiv:2503.06778v3 Announce Type: replace Abstract: Event annotation is important for identifying market changes, monitoring breaking news, and understanding sociological trends. Although expert annot

Practical exposure correction via compensation

HardwareDGX agent

arXiv:2212.14245v2 Announce Type: replace Abstract: In computer vision, correcting the exposure level is a fundamental task for enhancing the visual quality of observations with inappropriate lightnes

Towards interpretable AI with quantum annealing feature selection

ResearchDGX agent

arXiv:2604.25649v1 Announce Type: new Abstract: Deep learning models are used in critical applications, in which mistakes can have serious consequences. Therefore, it is crucial to understand how and

Use of What-if Scenarios to Help Explain Artificial Intelligence Models for Neonatal Health

ResearchDGX agent

arXiv:2410.09635v2 Announce Type: replace Abstract: Early detection of intrapartum risks enables timely interventions to prevent or mitigate adverse labor outcomes such as cerebral palsy. However, acc

VOYAGER: A Training Free Approach for Generating Diverse Datasets using LLMs

ResearchDGX agent

arXiv:2512.12072v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly being used to generate synthetic datasets for the evaluation and training of downstream models. Howeve

28 Apr 2026

A2DEPT: Large Language Model-Driven Automated Algorithm Design via Evolutionary Program Trees

ResearchDGX agent

arXiv:2604.24043v1 Announce Type: new Abstract: Designing heuristics for combinatorial optimization problems (COPs) is a fundamental yet challenging task that traditionally requires extensive domain e

BiMol-Diff: A Unified Diffusion Framework for Molecular Generation and Captioning

ResearchDGX agent

arXiv:2604.24089v1 Announce Type: new Abstract: Bridging molecular structures and natural language is essential for controllable design. Autoregressive models struggle with long-range dependencies, wh

BVI-Mamba: Video Enhancement Using a Visual State-Space Model for Low-Light and Underwater Environments

SafetyDGX agent

arXiv:2604.23655v1 Announce Type: new Abstract: Videos captured in low-light and underwater conditions often suffer from distortions such as noise, low contrast, color imbalance, and blur. These issue

Can Aha Moments Be Fake? Identifying True and Decorative Thinking Steps in Chain-of-Thought

Model ReleasesDGX agent

arXiv:2510.24941v3 Announce Type: replace Abstract: Large language models can generate long chain-of-thought (CoT) reasoning, but it remains unclear whether the verbalized steps reflect the models' in

Citation-Driven Multi-View Training for Patent Embeddings: QaECTER and Sophia-Bench

Model ReleasesDGX agent

arXiv:2604.22897v1 Announce Type: cross Abstract: Patent retrieval underpins critical decisions in innovation, examination, and IP strategy, yet progress has been hampered by the absence of benchmarks

Comparative Study of Weighted and Coupled Second- and Fourth-Order PDEs for Image Despeckling in Grayscale, Color, SAR, and Ultrasound

Model ReleasesDGX agent

arXiv:2604.23612v1 Announce Type: new Abstract: Partial Differential Equation (PDE)-based approaches have gained significant attention in image despeckling due to their strong capability to preserve s

DVPO: Distributional Value Modeling-based Policy Optimization for LLM Post-Training

SafetyDGX agent

arXiv:2512.03847v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has shown strong performance in LLM post-training, but real-world deployment often involves noisy or incomplete su

Enabling Transparent Cyber Threat Intelligence Combining Large Language Models and Domain Ontologies

AgentsDGX agent

arXiv:2509.00081v2 Announce Type: replace-cross Abstract: Effective Cyber Threat Intelligence (CTI) relies upon accurately structured and semantically enriched information extracted from cybersecurity

Generalizable Friction Coefficient Estimation via Material Embedding and Proxy Interaction Modeling

ResearchDGX agent

arXiv:2604.24188v1 Announce Type: new Abstract: Accurately estimating friction coefficients between arbitrary material pairs is critical for robotics, digital fabrication, and physics-based simulation

Gradient-Guided Exploration of Generative Model's Latent Space for Controlled Iris Image Augmentations

ApplicationsDGX agent

arXiv:2511.09749v2 Announce Type: replace Abstract: Developing reliable iris recognition and presentation attack detection methods requires diverse datasets that capture realistic variations in iris f

JudgeSense: A Benchmark for Prompt Sensitivity in LLM-as-a-Judge Systems

Model ReleasesDGX agent

arXiv:2604.23478v1 Announce Type: new Abstract: Large language models are increasingly deployed as automated judges for evaluating other models, yet the stability of their verdicts under semantically

Judging the Judges: A Systematic Evaluation of Bias Mitigation Strategies in LLM-as-a-Judge Pipelines

Model ReleasesDGX agent

arXiv:2604.23178v1 Announce Type: new Abstract: LLM-as-a-Judge has become the dominant paradigm for evaluating language model outputs, yet LLM judges exhibit systematic biases that compromise evaluati

MEASER: Malware embedding attacks on open-source LLMs

Model ReleasesDGX agent

arXiv:2510.10486v2 Announce Type: replace-cross Abstract: Open-source large language models (LLMs) have demonstrated considerable dominance over proprietary LLMs in resolving neural processing tasks,

Resource-Constrained UAV-Based Weed Detection for Site-Specific Management on Edge Devices

Local AiDGX agent

arXiv:2604.23442v1 Announce Type: new Abstract: Weeds compete with crops for light, water, and nutrients, reducing yield and crop quality. Efficient weed detection is essential for site-specific weed

Scheming Ability in LLM-to-LLM Strategic Interactions

Model ReleasesDGX agent

arXiv:2510.12826v2 Announce Type: replace-cross Abstract: As large language model (LLM) agents are deployed autonomously in diverse contexts, evaluating their capacity for strategic deception becomes

SPEAR-1: Scaling Beyond Robot Demonstrations via 3D Understanding

ResearchDGX agent

arXiv:2511.17411v2 Announce Type: replace-cross Abstract: Robotic Foundation Models (RFMs) hold great promise as generalist, end-to-end systems for robot control. Yet their ability to generalize acros

27 Apr 2026

Adversarial Co-Evolution of Malware and Detection Models: A Bilevel Optimization Perspective

ResearchDGX agent

arXiv:2604.22569v1 Announce Type: cross Abstract: Machine learning-based malware detectors are increasingly vulnerable to adversarial examples. Traditional defenses, such as one-shot adversarial train

False Feasibility in Variable Impedance MPC for Legged Locomotion

Model ReleasesDGX agent

arXiv:2604.22251v1 Announce Type: new Abstract: Variable impedance model predictive control (MPC) formulations that treat joint stiffness as an instantaneous decision variable operate on a feasible se

How LLMs Detect and Correct Their Own Errors: The Role of Internal Confidence Signals

Model ReleasesDGX agent

arXiv:2604.22271v1 Announce Type: new Abstract: Large language models can detect their own errors and sometimes correct them without external feedback, but the underlying mechanisms remain unknown. We

How Popsa used Amazon Nova to inspire customers with personalised title suggestions

Model ReleasesDGX agent

In this post, we share how we applied Amazon Bedrock and the Amazon Nova family of models to reimagine our Title Suggestion feature. By combining metadata, computer vision, and retrieval-augmented gen

microsoft/VibeVoice

Model ReleasesDGX agent

microsoft/VibeVoice VibeVoice is Microsoft's Whisper-style audio model for speech-to-text, MIT licensed and with speaker diarization built into the model. Microsoft released it on January 21st, 2026 b

RouteLMT: Learned Sample Routing for Hybrid LLM Translation Deployment

ResearchDGX agent

arXiv:2604.22520v1 Announce Type: new Abstract: Large Language Models (LLMs) have achieved remarkable performance in Machine Translation (MT), but deploying them at scale remains prohibitively expensi

Verbal Confidence Saturation in 3-9B Open-Weight Instruction-Tuned LLMs: A Pre-Registered Psychometric Validity Screen

ResearchDGX agent

arXiv:2604.22215v1 Announce Type: cross Abstract: Verbal confidence elicitation is widely used to extract uncertainty estimates from LLMs. We tested whether seven instruction-tuned open-weight models

25 Apr 2026

Why there is no cloud version for Qwen 3.6 27/35B?

Model ReleasesDGX agent

The Qwen 3.6-27B and 35B models are designed as open-weight models that developers can run locally on their own hardware without requiring cloud services. Alibaba released a separate cloud-only produc

24 Apr 2026

Attention-based multiple instance learning for predominant growth pattern prediction in lung adenocarcinoma wsi using foundation models

ResearchDGX agent

arXiv:2604.21530v1 Announce Type: cross Abstract: Lung adenocarcinoma (LUAD) grading depends on accurately identifying growth patterns, which are indicators of prognosis and can influence treatment de

Basic syntax from speech: Spontaneous concatenation in unsupervised deep neural networks

TutorialsDGX agent

arXiv:2305.01626v4 Announce Type: replace-cross Abstract: Computational models of syntax are predominantly text-based. Here we propose that the most basic first step in the evolution of syntax can be

Design, Modelling and Experimental Evaluation of a Tendon-driven Wrist Abduction-Adduction Mechanism for an upper limb exoskeleton

TutorialsDGX agent

arXiv:2604.20893v1 Announce Type: new Abstract: Wrist exoskeletons play a vital role in rehabilitation and assistive applications, yet conventional actuation mechanisms such as electric motors or pneu

Intent Laundering: AI Safety Datasets Are Not What They Seem

Model ReleasesDGX agent

arXiv:2602.16729v3 Announce Type: replace-cross Abstract: We systematically evaluate the quality of widely used adversarial safety datasets from two perspectives: in isolation and in practice. In isol

LiveVLM: Efficient Online Video Understanding via Streaming-Oriented KV Cache and Retrieval

Model ReleasesDGX agent

arXiv:2505.15269v2 Announce Type: replace Abstract: Recent developments in Video Large Language Models (Video LLMs) have enabled models to process hour-long videos and exhibit exceptional performance.

llm 0.31

Model ReleasesDGX agent

Release: llm 0.31 New GPT-5.5 OpenAI model: llm -m gpt-5.5. #1418 New option to set the text verbosity level for GPT-5+ OpenAI models: -o verbosity low. Values are low, medium, high. New option for se

Reasoning Primitives in Hybrid and Non-Hybrid LLMs

ApplicationsDGX agent

arXiv:2604.21454v1 Announce Type: cross Abstract: Reasoning in large language models is often treated as a monolithic capability, but its observed gains may arise from more basic operations. We study

SurgViVQA: Temporally-Grounded Video Question Answering for Surgical Scene Understanding

Model ReleasesDGX agent

arXiv:2511.03325v3 Announce Type: replace Abstract: Video Question Answering (VideoQA) in the surgical domain aims to enhance intraoperative understanding by enabling AI models to reason over temporal

VG-CoT: Towards Trustworthy Visual Reasoning via Grounded Chain-of-Thought

Model ReleasesDGX agent

arXiv:2604.21396v1 Announce Type: cross Abstract: The advancement of Large Vision-Language Models (LVLMs) requires precise local region-based reasoning that faithfully grounds the model's logic in act

Welcome DeepSeek V4 Pro Max https://huggingface.co/deepseek-ai/DeepSeek-V4-Pro

Model ReleasesDGX agent

DeepSeek V4 Pro Max is a large language model released by DeepSeek AI and made available on Hugging Face, representing an advancement in their model lineup. The announcement was made by Clem Delangue,

← Previous
1…239240241242243…1018
Next →