AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
All
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,280 results
Model Releases

Governed Reasoning for Institutional AI

DGX agent

arXiv:2604.10658v1 Announce Type: new Abstract: Institutional decisions -- regulatory compliance, clinical triage, prior authorization appeal -- require a different AI architecture than general-purpos

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

GPT vs Claude in a bomberman-style 1v1 game

DGX agent
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

A Reddit post on r/ChatGPT in which a user built or showcased a Bomberman-style 1v1 game pitting GPT (OpenAI) against Claude (Anthropic) as autonomous AI players, likely using their respective APIs to

model-releasesr-chatgpt
14 Apr 2026
Model Releases

Gradient-Variation Regret Bounds for Unconstrained Online Learning

DGX agent

arXiv:2604.11151v1 Announce Type: new Abstract: We develop parameter-free algorithms for unconstrained online learning with regret guarantees that scale with the gradient variation V_T(u) = sum_{t=2}^

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

Graph Retention Networks for Dynamic Graphs

DGX agent

arXiv:2411.11259v3 Announce Type: replace Abstract: In this paper, we propose Graph Retention Networks (GRNs) as a unified architecture for deep learning on dynamic graphs. The GRN extends the concept

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

Grid2Matrix: Revealing Digital Agnosia in Vision-Language Models

DGX agent

arXiv:2604.09687v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) excel on many multimodal reasoning benchmarks, but these evaluations often do not require an exhaustive readout of the i

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Grounded World Model for Semantically Generalizable Planning

DGX agent

arXiv:2604.11751v1 Announce Type: cross Abstract: In Model Predictive Control (MPC), world models predict the future outcomes of various action proposals, which are then scored to guide the selection

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

GroupRank: A Groupwise Paradigm for Effective and Efficient Passage Reranking with LLMs

DGX agent

arXiv:2511.11653v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have emerged as powerful tools for passage reranking in information retrieval, leveraging their superior reasonin

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

HDR 3D Gaussian Splatting via Luminance-Chromaticity Decomposition

DGX agent

arXiv:2511.12895v2 Announce Type: replace Abstract: High Dynamic Range (HDR) 3D reconstruction is pivotal for professional content creation in filmmaking and virtual production. Existing methods typic

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

HealthAdminBench: Evaluating Computer-Use Agents on Healthcare Administration Tasks

DGX agent

arXiv:2604.09937v1 Announce Type: new Abstract: Healthcare administration accounts for over $1 trillion in annual spending, making it a promising target for LLM-based computer-use agents (CUAs). While

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

HearthNet: Edge Multi-Agent Orchestration for Smart Homes

DGX agent

arXiv:2604.09618v1 Announce Type: cross Abstract: Smart-home users increasingly want to control their homes in natural language rather than assemble rules, dashboards, and API integrations by hand. At

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

HeceTokenizer: A Syllable-Based Tokenization Approach for Turkish Retrieval

DGX agent

arXiv:2604.10665v1 Announce Type: new Abstract: HeceTokenizer is a syllable-based tokenizer for Turkish that exploits the deterministic six-pattern phonological structure of the language to construct

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

[Help] Gemma 4 26B LoRA Training on 16GB VRAM: Loss decreases, but inference degenerates into loops (Masking vs. MoE?)

DGX agent

This Reddit thread discusses a user's experience attempting LoRA fine-tuning of the Gemma 4 26B-A4B model on a 16GB VRAM GPU, where training loss decreases normally but the resulting model degenerates

model-releasesr-ollama
14 Apr 2026
Model Releases

here’s a good application of harness permissions Programmatically Enforced Auto-Research: - auto-research loops usually expose a set of file…

DGX agent

here’s a good application of harness permissions Programmatically Enforced Auto-Research: - auto-research loops usually expose a set of files that the agent is allowed to edit to hill climb a metric/e

model-releasesharrison-chase--x
14 Apr 2026
Model Releases

Heterogeneous Connectivity in Sparse Networks: Fan-in Profiles, Gradient Hierarchy, and Topological Equilibria

DGX agent

arXiv:2604.10560v1 Announce Type: new Abstract: Profiled Sparse Networks (PSN) replace uniform connectivity with deterministic, heterogeneous fan-in profiles defined by continuous, nonlinear functions

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

HG-Lane: High-Fidelity Generation of Lane Scenes under Adverse Weather and Lighting Conditions without Re-annotation

DGX agent

arXiv:2603.10128v2 Announce Type: replace Abstract: Lane detection is a crucial task in autonomous driving, as it helps ensure the safe operation of vehicles. However, existing datasets such as CULane

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

HiEdit: Lifelong Model Editing with Hierarchical Reinforcement Learning

DGX agent

arXiv:2604.11214v1 Announce Type: new Abstract: Lifelong model editing (LME) aims to sequentially rectify outdated or inaccurate knowledge in deployed LLMs while minimizing side effects on unrelated i

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

HiPRAG: Hierarchical Process Rewards for Efficient Agentic Retrieval Augmented Generation

DGX agent

arXiv:2510.07794v2 Announce Type: replace-cross Abstract: Agentic RAG is a powerful technique for incorporating external information that LLMs lack, enabling better problem solving and question answer

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Hodoscope: Unsupervised Monitoring for AI Misbehaviors

DGX agent

arXiv:2604.11072v1 Announce Type: new Abstract: Existing approaches to monitoring AI agents rely on supervised evaluation: human-written rules or LLM-based judges that check for known failure modes. H

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

How Alignment Routes: Localizing, Scaling, and Controlling Policy Circuits in Language Models

DGX agent

arXiv:2604.04385v3 Announce Type: replace-cross Abstract: This paper localizes the policy routing mechanism in alignment-trained language models. An intermediate-layer attention gate reads detected co

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

How Controllable Are Large Language Models? A Unified Evaluation across Behavioral Granularities

DGX agent

arXiv:2603.02578v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly deployed in socially sensitive domains, yet their unpredictable behaviors, ranging from misalign

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

How Many Tries Does It Take? Iterative Self-Repair in LLM Code Generation Across Model Scales and Benchmarks

DGX agent

arXiv:2604.10508v1 Announce Type: cross Abstract: Large language models frequently fail to produce correct code on their first attempt, yet most benchmarks evaluate them in a single-shot setting. We i

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

How Robust Are Large Language Models for Clinical Numeracy? An Empirical Study on Numerical Reasoning Abilities in Clinical Contexts

DGX agent

arXiv:2604.11133v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly being explored for clinical question answering and decision support, yet safe deployment critically requir

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

How You Ask Matters! Adaptive RAG Robustness to Query Variations

DGX agent

arXiv:2604.10745v1 Announce Type: new Abstract: Adaptive Retrieval-Augmented Generation (RAG) promises accuracy and efficiency by dynamically triggering retrieval only when needed and is widely used i

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

HumanVBench: Probing Human-Centric Video Understanding in MLLMs with Automatically Synthesized Benchmarks

DGX agent

arXiv:2412.17574v3 Announce Type: replace-cross Abstract: Evaluating the nuanced human-centric video understanding capabilities of Multimodal Large Language Models (MLLMs) remains a great challenge, a

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

HumorGen: Cognitive Synergy for Humor Generation in Large Language Models via Persona-Based Distillation

DGX agent

arXiv:2604.09629v1 Announce Type: new Abstract: Humor generation poses a significant challenge for Large Language Models (LLMs), because their standard training objective - predicting the most likely

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Hunt Globally: Wide Search AI Agents for Drug Asset Scouting in Investing, Business Development, and Competitive Intelligence

DGX agent

arXiv:2602.15019v3 Announce Type: replace Abstract: Bio-pharmaceutical innovation has shifted: many new drug assets now originate outside the United States and are disclosed primarily via regional, no

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Hypergraph Neural Diffusion: A PDE-Inspired Framework for Hypergraph Message Passing

DGX agent

arXiv:2604.10955v1 Announce Type: new Abstract: Hypergraph neural networks (HGNNs) have shown remarkable potential in modeling high-order relationships that naturally arise in many real-world data dom

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

I actually cancelled my Claude Max subscription (well, downgraded to Pro, still need Deep Research) for Hermes with 1T+ parameter Chinese re…

DGX agent

Nous Research's Hermes model, a large-scale Chinese-trained model with over 1 trillion parameters, prompted at least one user to cancel or downgrade their Claude Max subscription in favor of it, retai

model-releasesnous-research--x
14 Apr 2026
Model Releases

I Can't Believe TTA Is Not Better: When Test-Time Augmentation Hurts Medical Image Classification

DGX agent

arXiv:2604.09697v1 Announce Type: cross Abstract: Test-time augmentation (TTA)--aggregating predictions over multiple augmented copies of a test input--is widely assumed to improve classification accu

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

I have a Macbook AIR M5 Base and I want to run an Agentic Coding program, similar to Claude Code or Codex. Besides the model, how do I do it? I've already tried with Ollama, VS Code, Opencode, and haven't been able to. (I'm not a developer, sorry)

DGX agent

This Reddit thread addresses a common challenge for non-developers trying to run a local agentic coding assistant on a MacBook Air M5: while tools like Ollama, VS Code, and OpenCode are the right piec

model-releasesr-ollama
14 Apr 2026
Model Releases

ICYMI -- last week we released `deepagents deploy`, the fastest way to take a highly capable, long running agent to production. agents are b…

DGX agent

ICYMI -- last week we released `deepagents deploy`, the fastest way to take a highly capable, long running agent to production. agents are becoming more and more standardized, and we're betting on thi

model-releasesharrison-chase--x
14 Apr 2026
Model Releases

Identifying Inductive Biases for Robot Co-Design

DGX agent

arXiv:2604.11768v1 Announce Type: new Abstract: Co-designing a robot's morphology and control can ensure synergistic interactions between them, prevalent in biological organisms. However, co-design is

model-releasesarxiv-cs-ro
14 Apr 2026
Model Releases

If an LLM Were a Character, Would It Know Its Own Story? Evaluating Lifelong Learning in LLMs

DGX agent

arXiv:2503.23514v2 Announce Type: replace-cross Abstract: Large language models (LLMs) can carry out human-like dialogue, but unlike humans, they are stateless due to the superposition property. Howev

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

I'm excited about voice as a UI layer for existing visual applications — where speech and screen update together. This goes well beyond voic…

DGX agent

I'm excited about voice as a UI layer for existing visual applications — where speech and screen update together. This goes well beyond voice-only use cases like call center automation. The barrier ha

model-releasesandrew-ng--x
14 Apr 2026
Model Releases

Immunizing 3D Gaussian Generative Models Against Unauthorized Fine-Tuning via Attribute-Space Traps

DGX agent

arXiv:2604.09688v1 Announce Type: new Abstract: Recent large-scale generative models enable high-quality 3D synthesis. However, the public accessibility of pre-trained weights introduces a critical vu

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

IMPACT: A Dataset for Multi-Granularity Human Procedural Action Understanding in Industrial Assembly

DGX agent

arXiv:2604.10409v1 Announce Type: cross Abstract: We introduce IMPACT, a synchronized five-view RGB-D dataset for deployment-oriented industrial procedural understanding, built around real assembly an

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Incentivizing Honesty among Competitors in Collaborative Learning and Optimization

DGX agent

arXiv:2305.16272v5 Announce Type: replace Abstract: Collaborative learning techniques have the potential to enable training machine learning models that are superior to models trained on a single enti

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

Infusing Theory of Mind into Socially Intelligent LLM Agents

DGX agent

arXiv:2509.22887v2 Announce Type: replace Abstract: Theory of Mind (ToM)-an understanding of the mental states of others-is a key aspect of human social intelligence, yet, chatbots and LLM-based socia

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

INSPATIO-WORLD: A Real-Time 4D World Simulator via Spatiotemporal Autoregressive Modeling

DGX agent

arXiv:2604.07209v2 Announce Type: replace Abstract: Building world models with spatial consistency and real-time interactivity remains a fundamental challenge in computer vision. Current video generat

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Integrating Semi-Supervised and Active Learning for Semantic Segmentation

DGX agent

arXiv:2501.19227v2 Announce Type: replace-cross Abstract: In this paper, we propose a novel active learning approach integrated with an improved semi-supervised learning framework to reduce the cost o

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Intent-aligned Formal Specification Synthesis via Traceable Refinement

DGX agent

arXiv:2604.10392v1 Announce Type: cross Abstract: Large language models are increasingly used to generate code from natural language, but ensuring correctness remains challenging. Formal verification

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Intersectional Sycophancy: How Perceived User Demographics Shape False Validation in Large Language Models

DGX agent

arXiv:2604.11609v1 Announce Type: new Abstract: Large language models exhibit sycophantic tendencies--validating incorrect user beliefs to appear agreeable. We investigate whether this behavior varies

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Investigating Bias and Fairness in Appearance-based Gaze Estimation

DGX agent

arXiv:2604.10707v1 Announce Type: new Abstract: While appearance-based gaze estimation has achieved significant improvements in accuracy and domain adaptation, the fairness of these systems across dif

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Investigating Vaccine Buyer's Remorse: Post-Vaccination Decision Regret in COVID-19 Social Media Using Politically Diverse Human Annotation

DGX agent

arXiv:2604.09626v1 Announce Type: cross Abstract: A significant gap exists in datasets regarding post-COVID-19 vaccination experiences, particularly ``vaccine buyer's remorse''. Understanding the prev

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

Is There Knowledge Left to Extract? Evidence of Fragility in Medically Fine-Tuned Vision-Language Models

DGX agent

arXiv:2604.09841v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly adapted through domain-specific fine-tuning, yet it remains unclear whether this improves reasoning bey

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

It is going to be like what happened in coding: as soon as models crossed a certain threshold (Opus 4.5, GPT-5.2, Gemini 3), suddenly Claude…

DGX agent

It is going to be like what happened in coding: as soon as models crossed a certain threshold (Opus 4.5, GPT-5.2, Gemini 3), suddenly Claude Code & Codex were viable. Before that, it was all about cod

model-releasesethan-mollick--x
14 Apr 2026
Model Releases

It looks like everyone is finally catching up with the fact that agent sessions in CLI mode can only get you so far. It makes sense that the…

DGX agent

It looks like everyone is finally catching up with the fact that agent sessions in CLI mode can only get you so far. It makes sense that the new Codex app, Cursor, and Claude Code (desktop) feel and l

model-releasesdair-ai--x
14 Apr 2026
Model Releases

ITIScore: An Image-to-Text-to-Image Rating Framework for the Image Captioning Ability of MLLMs

DGX agent

arXiv:2604.03765v2 Announce Type: replace Abstract: Recent advances in multimodal large language models (MLLMs) have greatly improved image understanding and captioning capabilities. However, existing

model-releasesarxiv-cs-cv
14 Apr 2026
← Previous
1…440441442443444…465
Next →