AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,272 results
14 Apr 2026

HealthAdminBench: Evaluating Computer-Use Agents on Healthcare Administration Tasks

Model ReleasesDGX agent

arXiv:2604.09937v1 Announce Type: new Abstract: Healthcare administration accounts for over $1 trillion in annual spending, making it a promising target for LLM-based computer-use agents (CUAs). While

HearthNet: Edge Multi-Agent Orchestration for Smart Homes

Model ReleasesDGX agent

arXiv:2604.09618v1 Announce Type: cross Abstract: Smart-home users increasingly want to control their homes in natural language rather than assemble rules, dashboards, and API integrations by hand. At

HeceTokenizer: A Syllable-Based Tokenization Approach for Turkish Retrieval

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.10665v1 Announce Type: new Abstract: HeceTokenizer is a syllable-based tokenizer for Turkish that exploits the deterministic six-pattern phonological structure of the language to construct

[Help] Gemma 4 26B LoRA Training on 16GB VRAM: Loss decreases, but inference degenerates into loops (Masking vs. MoE?)

Model ReleasesDGX agent

This Reddit thread discusses a user's experience attempting LoRA fine-tuning of the Gemma 4 26B-A4B model on a 16GB VRAM GPU, where training loss decreases normally but the resulting model degenerates

here’s a good application of harness permissions Programmatically Enforced Auto-Research: - auto-research loops usually expose a set of file…

Model ReleasesDGX agent

here’s a good application of harness permissions Programmatically Enforced Auto-Research: - auto-research loops usually expose a set of files that the agent is allowed to edit to hill climb a metric/e

Heterogeneous Connectivity in Sparse Networks: Fan-in Profiles, Gradient Hierarchy, and Topological Equilibria

Model ReleasesDGX agent

arXiv:2604.10560v1 Announce Type: new Abstract: Profiled Sparse Networks (PSN) replace uniform connectivity with deterministic, heterogeneous fan-in profiles defined by continuous, nonlinear functions

HG-Lane: High-Fidelity Generation of Lane Scenes under Adverse Weather and Lighting Conditions without Re-annotation

Model ReleasesDGX agent

arXiv:2603.10128v2 Announce Type: replace Abstract: Lane detection is a crucial task in autonomous driving, as it helps ensure the safe operation of vehicles. However, existing datasets such as CULane

HiEdit: Lifelong Model Editing with Hierarchical Reinforcement Learning

Model ReleasesDGX agent

arXiv:2604.11214v1 Announce Type: new Abstract: Lifelong model editing (LME) aims to sequentially rectify outdated or inaccurate knowledge in deployed LLMs while minimizing side effects on unrelated i

HiPRAG: Hierarchical Process Rewards for Efficient Agentic Retrieval Augmented Generation

Model ReleasesDGX agent

arXiv:2510.07794v2 Announce Type: replace-cross Abstract: Agentic RAG is a powerful technique for incorporating external information that LLMs lack, enabling better problem solving and question answer

Hodoscope: Unsupervised Monitoring for AI Misbehaviors

Model ReleasesDGX agent

arXiv:2604.11072v1 Announce Type: new Abstract: Existing approaches to monitoring AI agents rely on supervised evaluation: human-written rules or LLM-based judges that check for known failure modes. H

How Alignment Routes: Localizing, Scaling, and Controlling Policy Circuits in Language Models

Model ReleasesDGX agent

arXiv:2604.04385v3 Announce Type: replace-cross Abstract: This paper localizes the policy routing mechanism in alignment-trained language models. An intermediate-layer attention gate reads detected co

How Controllable Are Large Language Models? A Unified Evaluation across Behavioral Granularities

Model ReleasesDGX agent

arXiv:2603.02578v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly deployed in socially sensitive domains, yet their unpredictable behaviors, ranging from misalign

How Many Tries Does It Take? Iterative Self-Repair in LLM Code Generation Across Model Scales and Benchmarks

Model ReleasesDGX agent

arXiv:2604.10508v1 Announce Type: cross Abstract: Large language models frequently fail to produce correct code on their first attempt, yet most benchmarks evaluate them in a single-shot setting. We i

How Robust Are Large Language Models for Clinical Numeracy? An Empirical Study on Numerical Reasoning Abilities in Clinical Contexts

Model ReleasesDGX agent

arXiv:2604.11133v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly being explored for clinical question answering and decision support, yet safe deployment critically requir

How You Ask Matters! Adaptive RAG Robustness to Query Variations

Model ReleasesDGX agent

arXiv:2604.10745v1 Announce Type: new Abstract: Adaptive Retrieval-Augmented Generation (RAG) promises accuracy and efficiency by dynamically triggering retrieval only when needed and is widely used i

HumanVBench: Probing Human-Centric Video Understanding in MLLMs with Automatically Synthesized Benchmarks

Model ReleasesDGX agent

arXiv:2412.17574v3 Announce Type: replace-cross Abstract: Evaluating the nuanced human-centric video understanding capabilities of Multimodal Large Language Models (MLLMs) remains a great challenge, a

HumorGen: Cognitive Synergy for Humor Generation in Large Language Models via Persona-Based Distillation

Model ReleasesDGX agent

arXiv:2604.09629v1 Announce Type: new Abstract: Humor generation poses a significant challenge for Large Language Models (LLMs), because their standard training objective - predicting the most likely

Hunt Globally: Wide Search AI Agents for Drug Asset Scouting in Investing, Business Development, and Competitive Intelligence

Model ReleasesDGX agent

arXiv:2602.15019v3 Announce Type: replace Abstract: Bio-pharmaceutical innovation has shifted: many new drug assets now originate outside the United States and are disclosed primarily via regional, no

Hypergraph Neural Diffusion: A PDE-Inspired Framework for Hypergraph Message Passing

Model ReleasesDGX agent

arXiv:2604.10955v1 Announce Type: new Abstract: Hypergraph neural networks (HGNNs) have shown remarkable potential in modeling high-order relationships that naturally arise in many real-world data dom

I actually cancelled my Claude Max subscription (well, downgraded to Pro, still need Deep Research) for Hermes with 1T+ parameter Chinese re…

Model ReleasesDGX agent

Nous Research's Hermes model, a large-scale Chinese-trained model with over 1 trillion parameters, prompted at least one user to cancel or downgrade their Claude Max subscription in favor of it, retai

I Can't Believe TTA Is Not Better: When Test-Time Augmentation Hurts Medical Image Classification

Model ReleasesDGX agent

arXiv:2604.09697v1 Announce Type: cross Abstract: Test-time augmentation (TTA)--aggregating predictions over multiple augmented copies of a test input--is widely assumed to improve classification accu

I have a Macbook AIR M5 Base and I want to run an Agentic Coding program, similar to Claude Code or Codex. Besides the model, how do I do it? I've already tried with Ollama, VS Code, Opencode, and haven't been able to. (I'm not a developer, sorry)

Model ReleasesDGX agent

This Reddit thread addresses a common challenge for non-developers trying to run a local agentic coding assistant on a MacBook Air M5: while tools like Ollama, VS Code, and OpenCode are the right piec

ICYMI -- last week we released `deepagents deploy`, the fastest way to take a highly capable, long running agent to production. agents are b…

Model ReleasesDGX agent

ICYMI -- last week we released `deepagents deploy`, the fastest way to take a highly capable, long running agent to production. agents are becoming more and more standardized, and we're betting on thi

Identifying Inductive Biases for Robot Co-Design

Model ReleasesDGX agent

arXiv:2604.11768v1 Announce Type: new Abstract: Co-designing a robot's morphology and control can ensure synergistic interactions between them, prevalent in biological organisms. However, co-design is

If an LLM Were a Character, Would It Know Its Own Story? Evaluating Lifelong Learning in LLMs

Model ReleasesDGX agent

arXiv:2503.23514v2 Announce Type: replace-cross Abstract: Large language models (LLMs) can carry out human-like dialogue, but unlike humans, they are stateless due to the superposition property. Howev

I'm excited about voice as a UI layer for existing visual applications — where speech and screen update together. This goes well beyond voic…

Model ReleasesDGX agent

I'm excited about voice as a UI layer for existing visual applications — where speech and screen update together. This goes well beyond voice-only use cases like call center automation. The barrier ha

Immunizing 3D Gaussian Generative Models Against Unauthorized Fine-Tuning via Attribute-Space Traps

Model ReleasesDGX agent

arXiv:2604.09688v1 Announce Type: new Abstract: Recent large-scale generative models enable high-quality 3D synthesis. However, the public accessibility of pre-trained weights introduces a critical vu

IMPACT: A Dataset for Multi-Granularity Human Procedural Action Understanding in Industrial Assembly

Model ReleasesDGX agent

arXiv:2604.10409v1 Announce Type: cross Abstract: We introduce IMPACT, a synchronized five-view RGB-D dataset for deployment-oriented industrial procedural understanding, built around real assembly an

Incentivizing Honesty among Competitors in Collaborative Learning and Optimization

Model ReleasesDGX agent

arXiv:2305.16272v5 Announce Type: replace Abstract: Collaborative learning techniques have the potential to enable training machine learning models that are superior to models trained on a single enti

Infusing Theory of Mind into Socially Intelligent LLM Agents

Model ReleasesDGX agent

arXiv:2509.22887v2 Announce Type: replace Abstract: Theory of Mind (ToM)-an understanding of the mental states of others-is a key aspect of human social intelligence, yet, chatbots and LLM-based socia

INSPATIO-WORLD: A Real-Time 4D World Simulator via Spatiotemporal Autoregressive Modeling

Model ReleasesDGX agent

arXiv:2604.07209v2 Announce Type: replace Abstract: Building world models with spatial consistency and real-time interactivity remains a fundamental challenge in computer vision. Current video generat

Integrating Semi-Supervised and Active Learning for Semantic Segmentation

Model ReleasesDGX agent

arXiv:2501.19227v2 Announce Type: replace-cross Abstract: In this paper, we propose a novel active learning approach integrated with an improved semi-supervised learning framework to reduce the cost o

Intent-aligned Formal Specification Synthesis via Traceable Refinement

Model ReleasesDGX agent

arXiv:2604.10392v1 Announce Type: cross Abstract: Large language models are increasingly used to generate code from natural language, but ensuring correctness remains challenging. Formal verification

Intersectional Sycophancy: How Perceived User Demographics Shape False Validation in Large Language Models

Model ReleasesDGX agent

arXiv:2604.11609v1 Announce Type: new Abstract: Large language models exhibit sycophantic tendencies--validating incorrect user beliefs to appear agreeable. We investigate whether this behavior varies

Investigating Bias and Fairness in Appearance-based Gaze Estimation

Model ReleasesDGX agent

arXiv:2604.10707v1 Announce Type: new Abstract: While appearance-based gaze estimation has achieved significant improvements in accuracy and domain adaptation, the fairness of these systems across dif

Investigating Vaccine Buyer's Remorse: Post-Vaccination Decision Regret in COVID-19 Social Media Using Politically Diverse Human Annotation

Model ReleasesDGX agent

arXiv:2604.09626v1 Announce Type: cross Abstract: A significant gap exists in datasets regarding post-COVID-19 vaccination experiences, particularly ``vaccine buyer's remorse''. Understanding the prev

Is There Knowledge Left to Extract? Evidence of Fragility in Medically Fine-Tuned Vision-Language Models

Model ReleasesDGX agent

arXiv:2604.09841v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly adapted through domain-specific fine-tuning, yet it remains unclear whether this improves reasoning bey

It is going to be like what happened in coding: as soon as models crossed a certain threshold (Opus 4.5, GPT-5.2, Gemini 3), suddenly Claude…

Model ReleasesDGX agent

It is going to be like what happened in coding: as soon as models crossed a certain threshold (Opus 4.5, GPT-5.2, Gemini 3), suddenly Claude Code & Codex were viable. Before that, it was all about cod

It looks like everyone is finally catching up with the fact that agent sessions in CLI mode can only get you so far. It makes sense that the…

Model ReleasesDGX agent

It looks like everyone is finally catching up with the fact that agent sessions in CLI mode can only get you so far. It makes sense that the new Codex app, Cursor, and Claude Code (desktop) feel and l

ITIScore: An Image-to-Text-to-Image Rating Framework for the Image Captioning Ability of MLLMs

Model ReleasesDGX agent

arXiv:2604.03765v2 Announce Type: replace Abstract: Recent advances in multimodal large language models (MLLMs) have greatly improved image understanding and captioning capabilities. However, existing

Japanese tech giants launch joint venture targeting physical AI for robots and machines

Model ReleasesDGX agent

Japanese technology giants SoftBank Group Corp., Sony Corp. and NEC Corp. are teaming up with Honda Motor Co., Ltd. on a new artificial intelligence joint venture that has a single goal: to build a tr

K-Way Energy Probes for Metacognition Reduce to Softmax in Discriminative Predictive Coding Networks

Model ReleasesDGX agent

arXiv:2604.11011v1 Announce Type: cross Abstract: We present this as a negative result with an explanatory mechanism, not as a formal upper bound. Predictive coding networks (PCNs) admit a K-way energ

Knowing What to Stress: A Discourse-Conditioned Text-to-Speech Benchmark

Model ReleasesDGX agent

arXiv:2604.10580v1 Announce Type: new Abstract: Spoken meaning often depends not only on what is said, but also on which word is emphasized. The same sentence can convey correction, contrast, or clari

Knowledge Integration in Differentiable Models: A Comparative Study of Data-Driven, Soft-Constrained, and Hard-Constrained Paradigms for Identification and Control of the Single Machine Infinite Bus System

Model ReleasesDGX agent

arXiv:2602.09667v2 Announce Type: replace Abstract: Integrating domain knowledge into neural networks is a central challenge in scientific machine learning. Three paradigms have emerged -- data-driven

ks-pret-5m: a 5 million word, 12 million token kashmiri pretraining dataset

Model ReleasesDGX agent

arXiv:2604.11066v1 Announce Type: new Abstract: We present KS-PRET-5M, the largest publicly available pretraining dataset for the Kashmiri language, comprising 5,090,244 (5.09M) words, 27,692,959 (27.

LABBench2: An Improved Benchmark for AI Systems Performing Biology Research

Model ReleasesDGX agent

arXiv:2604.09554v1 Announce Type: new Abstract: Optimism for accelerating scientific discovery with AI continues to grow. Current applications of AI in scientific research range from training dedicate

LaMI: Augmenting Large Language Models via Late Multi-Image Fusion

Model ReleasesDGX agent

arXiv:2406.13621v2 Announce Type: replace Abstract: Commonsense reasoning often requires both textual and visual knowledge, yet Large Language Models (LLMs) trained solely on text lack visual groundin

Language Prompt vs. Image Enhancement: Boosting Object Detection With CLIP in Hazy Environments

Model ReleasesDGX agent

arXiv:2604.10637v1 Announce Type: new Abstract: Object detection in hazy environments is challenging because degraded objects are nearly invisible and their semantics are weakened by environmental noi

Large Language Models Can Help Mitigate Barren Plateaus in Quantum Neural Networks

Model ReleasesDGX agent

arXiv:2502.13166v3 Announce Type: replace-cross Abstract: In the era of noisy intermediate-scale quantum (NISQ) computing, Quantum Neural Networks (QNNs) have emerged as a promising approach for vario

LARY: A Latent Action Representation Yielding Benchmark for Generalizable Vision-to-Action Alignment

Model ReleasesDGX agent

arXiv:2604.11689v1 Announce Type: new Abstract: While the shortage of explicit action data limits Vision-Language-Action (VLA) models, human action videos offer a scalable yet unlabeled data source. A

LAST: Leveraging Tools as Hints to Enhance Spatial Reasoning for Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2604.09712v1 Announce Type: cross Abstract: Spatial reasoning is a cornerstone capability for intelligent systems to perceive and interact with the physical world. However, multimodal large lang

LEADER: Learning Reliable Local-to-Global Correspondences for LiDAR Relocalization

Model ReleasesDGX agent

arXiv:2604.11355v1 Announce Type: new Abstract: LiDAR relocalization has attracted increasing attention as it can deliver accurate 6-DoF pose estimation in complex 3D environments. Recent learning-bas

Learning Racket-Ball Bounce Dynamics Across Diverse Rubbers for Robotic Table Tennis

Model ReleasesDGX agent

arXiv:2604.11349v1 Announce Type: new Abstract: Accurate dynamic models for racket-ball bounces are essential for reliable control in robotic table tennis. Existing models typically assume simple line

Learning Robustness at Test-Time from a Non-Robust Teacher

Model ReleasesDGX agent

arXiv:2604.11590v1 Announce Type: new Abstract: Nowadays, pretrained models are increasingly used as general-purpose backbones and adapted at test-time to downstream environments where target data are

Learning to Adapt: In-Context Learning Beyond Stationarity

Model ReleasesDGX agent

arXiv:2604.10946v1 Announce Type: new Abstract: Transformer models have become foundational across a wide range of scientific and engineering domains due to their strong empirical performance. A key c

Learning to Play Piano in the Real World

Model ReleasesDGX agent

arXiv:2503.15481v3 Announce Type: replace-cross Abstract: Towards the grand challenge of achieving human-level manipulation in robots, playing piano is a compelling testbed that requires strategic, pr

Learning What's Real: Disentangling Signal and Measurement Artifacts in Multi-Sensor Data, with Applications to Astrophysics

Model ReleasesDGX agent

arXiv:2604.09787v1 Announce Type: cross Abstract: Data collected from the physical world is always a combination of multiple sources: an underlying signal from the physical process of interest and a s

Learning World Models for Interactive Video Generation

Model ReleasesDGX agent

arXiv:2505.21996v3 Announce Type: replace-cross Abstract: Foundational world models must be both interactive and preserve spatiotemporal coherence for effective future planning with action choices. Ho

let's go open-source and local models!

Model ReleasesDGX agent

let's go open-source and local models! Uber's CTO told @LauraBratton5 that AI coding tools—particularly Anthropic’s Claude Code—has already maxed out its 2026 AI budget 📈 “I'm back to the drawing boar

LIDARLearn: A Unified Deep Learning Library for 3D Point Cloud Classification, Segmentation, and Self-Supervised Representation Learning

Model ReleasesDGX agent

arXiv:2604.10780v1 Announce Type: new Abstract: Three-dimensional (3D) point cloud analysis has become central to applications ranging from autonomous driving and robotics to forestry and ecological m

← Previous
1…351352353354355…372
Next →