AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
All
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,284 results
Model Releases

MemGround: Long-Term Memory Evaluation Kit for Large Language Models in Gamified Scenarios

DGX agent

arXiv:2604.14158v1 Announce Type: new Abstract: Current evaluations of long-term memory in LLMs are fundamentally static. By fixating on simple retrieval and short-context inference, they neglect the

model-releasesarxiv-cs-cl
17 Apr 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

MetaDent: Labeling Clinical Images for Vision-Language Models in Dentistry

DGX agent

arXiv:2604.14866v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have demonstrated significant potential in medical image analysis, yet their application in intraoral photography remains

model-releasesarxiv-cs-cv
17 Apr 2026
Model Releases

Mitigating LLM biases toward spurious social contexts using direct preference optimization

DGX agent

arXiv:2604.02585v2 Announce Type: replace-cross Abstract: LLMs are increasingly used for high-stakes decision-making, yet their sensitivity to spurious contextual information can introduce harmful bia

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

MixAtlas: Uncertainty-aware Data Mixture Optimization for Multimodal LLM Midtraining

DGX agent

arXiv:2604.14198v1 Announce Type: cross Abstract: Domain reweighting can improve sample efficiency and downstream generalization, but data-mixture optimization for multimodal midtraining remains large

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

MM-WebAgent: A Hierarchical Multimodal Web Agent for Webpage Generation

DGX agent

arXiv:2604.15309v1 Announce Type: cross Abstract: The rapid progress of Artificial Intelligence Generated Content (AIGC) tools enables images, videos, and visualizations to be created on demand for we

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

Modular Continual Learning via Zero-Leakage Reconstruction Routing and Autonomous Task Discovery

DGX agent

arXiv:2604.14375v1 Announce Type: new Abstract: Catastrophic forgetting remains a primary hurdle in sequential task learning for artificial neural networks. We propose a silicon-native modular archite

model-releasesarxiv-cs-lg
17 Apr 2026
Model Releases

ModuSeg: Decoupling Object Discovery and Semantic Retrieval for Training-Free Weakly Supervised Segmentation

DGX agent

arXiv:2604.07021v2 Announce Type: replace Abstract: Weakly supervised semantic segmentation aims to achieve pixel-level predictions using image-level labels. Existing methods typically entangle semant

model-releasesarxiv-cs-cv
17 Apr 2026
Model Releases

Multigrain-aware Semantic Prototype Scanning and Tri-Token Prompt Learning Embraced High-Order RWKV for Pan-Sharpening

DGX agent

arXiv:2604.14622v1 Announce Type: new Abstract: In this work, we propose a Multigrain-aware Semantic Prototype Scanning paradigm for pan-sharpening, built upon a high-order RWKV architecture and a tri

model-releasesarxiv-cs-cv
17 Apr 2026
Model Releases

Neuro-Oracle: A Trajectory-Aware Agentic RAG Framework for Interpretable Epilepsy Surgical Prognosis

DGX agent

arXiv:2604.14216v1 Announce Type: cross Abstract: Predicting post-surgical seizure outcomes in pharmacoresistant epilepsy is a clinical challenge. Conventional deep-learning approaches operate on stat

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

OmniGCD: Abstracting Generalized Category Discovery for Modality Agnosticism

DGX agent

arXiv:2604.14762v1 Announce Type: new Abstract: Generalized Category Discovery (GCD) challenges methods to identify known and novel classes using partially labeled data, mirroring human category learn

model-releasesarxiv-cs-cv
17 Apr 2026
Model Releases

Open-Set Vein Biometric Recognition with Deep Metric Learning

DGX agent

arXiv:2604.14874v1 Announce Type: new Abstract: Most state-of-the-art vein recognition methods rely on closed-set classification, which inherently limits their scalability and prevents the adaptive en

model-releasesarxiv-cs-cv
17 Apr 2026
Model Releases

OpenAI ratchets up Codex’s agentic capabilities to rival Claude Code

DGX agent

OpenAI Group PBC today announced a major revamp of its artificial intelligence coding tool Codex, giving it a number of new “agentic” capabilities that enable more complex task automation. The ChatGPT

model-releasessiliconangle
17 Apr 2026
Model Releases

OpenMobile: Building Open Mobile Agents with Task and Trajectory Synthesis

DGX agent

arXiv:2604.15093v1 Announce Type: cross Abstract: Mobile agents powered by vision-language models have demonstrated impressive capabilities in automating mobile tasks, with recent leading models achie

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

Orak: A Foundational Benchmark for Training and Evaluating LLM Agents on Diverse Video Games

DGX agent

arXiv:2506.03610v3 Announce Type: replace Abstract: Large Language Model (LLM) agents are reshaping the game industry, by enabling more intelligent and human-preferable characters. Yet, current game b

model-releasesarxiv-cs-ai
17 Apr 2026
Model Releases

Our virtual hackathon is back! Join us for a week of building with Opus 4.7 alongside developers from around the world. The Claude Code team…

DGX agent

Our virtual hackathon is back! Join us for a week of building with Opus 4.7 alongside developers from around the world. The Claude Code team will be in the room all week, with a prize pool of $100K in

model-releasesboris-cherny--x
17 Apr 2026
Model Releases

Pangu-ACE: Adaptive Cascaded Experts for Educational Response Generation on EduBench

DGX agent

arXiv:2604.14828v1 Announce Type: new Abstract: Educational assistants should spend more computation only when the task needs it. This paper rewrites our earlier draft around the system that was actua

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

Parameter estimation for land-surface models using Neural Physics

DGX agent

arXiv:2505.02979v3 Announce Type: replace-cross Abstract: We propose a novel inverse-modelling approach which estimates the parameters of a simple land-surface model (LSM) by assimilating data into a

model-releasesarxiv-cs-lg
17 Apr 2026
Model Releases

PeerPrism: Peer Evaluation Expertise vs Review-writing AI

DGX agent

arXiv:2604.14513v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used in scientific peer review, assisting with drafting, rewriting, expansion, and refinement. However, ex

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

Physical Intelligence says its new model, π0.7, can direct robots on tasks they weren't trained on, an 'early sign' of generalization, surprising researchers (Connie Loizos/TechCrunch)

DGX agent

Connie Loizos / TechCrunch: Physical Intelligence says its new model, π0.7, can direct robots on tasks they weren't trained on, an “early sign” of generalization, surprising researchers — Physical Int

model-releasestechmeme
17 Apr 2026
Model Releases

Physically-Induced Atmospheric Adversarial Perturbations: Enhancing Transferability and Robustness in Remote Sensing Image Classification

DGX agent

arXiv:2604.14643v1 Announce Type: new Abstract: Adversarial attacks pose a severe threat to the reliability of deep learning models in remote sensing (RS) image classification. Most existing methods r

model-releasesarxiv-cs-cv
17 Apr 2026
Model Releases

PolyBench: Benchmarking LLM Forecasting and Trading Capabilities on Live Prediction Market Data

DGX agent

arXiv:2604.14199v1 Announce Type: cross Abstract: Predicting real-world events from live market signals demands systems that fuse qualitative news with quantitative order-book dynamics under strict te

model-releasesarxiv-cs-lg
17 Apr 2026
Model Releases

POP: Prefill-Only Pruning for Efficient Large Model Inference

DGX agent

arXiv:2602.03295v2 Announce Type: replace Abstract: Large Language Models (LLMs) and Vision-Language Models (VLMs) have demonstrated remarkable capabilities. However, their deployment is hindered by s

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

PortraitCraft: A Benchmark for Portrait Composition Understanding and Generation

DGX agent

arXiv:2604.03611v2 Announce Type: replace Abstract: Portrait composition plays a central role in portrait aesthetics and visual communication, yet existing datasets and benchmarks mainly focus on coar

model-releasesarxiv-cs-cv
17 Apr 2026
Model Releases

Prism: Symbolic Superoptimization of Tensor Programs

DGX agent

arXiv:2604.15272v1 Announce Type: cross Abstract: This paper presents Prism, the first symbolic superoptimizer for tensor programs. The key idea is sGraph, a symbolic, hierarchical representation that

model-releasesarxiv-cs-lg
17 Apr 2026
Model Releases

Prompt-Guided Image Editing with Masked Logit Nudging in Visual Autoregressive Models

DGX agent

arXiv:2604.14591v1 Announce Type: new Abstract: We address the problem of prompt-guided image editing in visual autoregressive models. Given a source image and a target text prompt, we aim to modify t

model-releasesarxiv-cs-cv
17 Apr 2026
Model Releases

Prompt Optimization Is a Coin Flip: Diagnosing When It Helps in Compound AI Systems

DGX agent

arXiv:2604.14585v1 Announce Type: cross Abstract: Prompt optimization in compound AI systems is statistically indistinguishable from a coin flip: across 72 optimization runs on Claude Haiku (6 methods

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

ProRank: Prompt Warmup via Reinforcement Learning for Small Language Models Reranking

DGX agent

arXiv:2506.03487v3 Announce Type: replace-cross Abstract: Reranking is fundamental to information retrieval and retrieval-augmented generation, with recent Large Language Models (LLMs) significantly a

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

Purging the Gray Zone: Latent-Geometric Denoising for Precise Knowledge Boundary Awareness

DGX agent

arXiv:2604.14324v1 Announce Type: new Abstract: Large language models (LLMs) often exhibit hallucinations due to their inability to accurately perceive their own knowledge boundaries. Existing abstent

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

QuantCode-Bench: A Benchmark for Evaluating the Ability of Large Language Models to Generate Executable Algorithmic Trading Strategies

DGX agent

arXiv:2604.15151v1 Announce Type: new Abstract: Large language models have demonstrated strong performance on general-purpose programming tasks, yet their ability to generate executable algorithmic tr

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

Query pipeline optimization for cancer patient question answering systems

DGX agent

arXiv:2412.14751v2 Announce Type: replace Abstract: Retrieval-augmented generation (RAG) mitigates hallucination in Large Language Models (LLMs) by using query pipelines to retrieve relevant external

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

Regret Tail Characterization of Optimal Bandit Algorithms with Generic Rewards

DGX agent

arXiv:2604.14876v1 Announce Type: cross Abstract: We study the tail behavior of regret in stochastic multi-armed bandits for algorithms that are asymptotically optimal in expectation. While minimizing

model-releasesarxiv-cs-lg
17 Apr 2026
Model Releases

RELOAD: A Robust and Efficient Learned Query Optimizer for Database Systems

DGX agent

arXiv:2604.14725v1 Announce Type: cross Abstract: Recent advances in query optimization have shifted from traditional rule-based and cost-based techniques towards machine learning-driven approaches. A

model-releasesarxiv-cs-lg
17 Apr 2026
Model Releases

Rethinking Patient Education as Multi-turn Multi-modal Interaction

DGX agent

arXiv:2604.14656v1 Announce Type: cross Abstract: Most medical multimodal benchmarks focus on static tasks such as image question answering, report generation, and plain-language rewriting. Patient ed

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

Retrieve, Then Classify: Corpus-Grounded Automation of Clinical Value Set Authoring

DGX agent

arXiv:2604.14616v1 Announce Type: new Abstract: Clinical value set authoring -- the task of identifying all codes in a standardized vocabulary that define a clinical concept -- is a recurring bottlene

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

ReviewGrounder: Improving Review Substantiveness with Rubric-Guided, Tool-Integrated Agents

DGX agent

arXiv:2604.14261v1 Announce Type: new Abstract: The rapid rise in AI conference submissions has driven increasing exploration of large language models (LLMs) for peer review support. However, LLM-base

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

Right at My Level: A Unified Multilingual Framework for Proficiency-Aware Text Simplification

DGX agent

arXiv:2604.05302v2 Announce Type: replace Abstract: Text simplification supports second language (L2) learning by providing comprehensible input, consistent with the Input Hypothesis. However, constru

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

SafeHarness: Lifecycle-Integrated Security Architecture for LLM-based Agent Deployment

DGX agent

arXiv:2604.13630v1 Announce Type: cross Abstract: The performance of large language model (LLM) agents depends critically on the execution harness, the system layer that orchestrates tool use, context

model-releasesarxiv-cs-ai
17 Apr 2026
Model Releases

SAGE Celer 2.6 Technical Card

DGX agent

arXiv:2604.14168v1 Announce Type: new Abstract: We introduce SAGE Celer 2.6, the latest in our line of general-purpose Celer models from SAGEA. Celer 2.6 is available in 5B, 10B, and 27B parameter siz

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

SAGE: Sign-Adaptive Gradient for Memory-Efficient LLM Optimization

DGX agent

arXiv:2604.07663v2 Announce Type: replace Abstract: The AdamW optimizer, while standard for LLM pretraining, is a critical memory bottleneck, consuming optimizer states equivalent to twice the model's

model-releasesarxiv-cs-lg
17 Apr 2026
Model Releases

SAQ: Stabilizer-Aware Quantum Error Correction Decoder

DGX agent

arXiv:2512.08914v2 Announce Type: replace-cross Abstract: Quantum Error Correction (QEC) decoding faces a fundamental accuracy-efficiency tradeoff. Classical methods like Minimum Weight Perfect Matchi

model-releasesarxiv-cs-ai
17 Apr 2026
Model Releases

Schema Key Wording as an Instruction Channel in Structured Generation under Constrained Decoding

DGX agent

arXiv:2604.14862v1 Announce Type: new Abstract: Constrained decoding has been widely adopted for structured generation with large language models (LLMs), ensuring that outputs satisfy predefined forma

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

Secure and Privacy-Preserving Vertical Federated Learning

DGX agent

arXiv:2604.13474v1 Announce Type: cross Abstract: We propose a novel end-to-end privacy-preserving framework, instantiated by three efficient protocols for different deployment scenarios, covering bot

model-releasesarxiv-cs-ai
17 Apr 2026
Model Releases

SelfGrader: Stable Jailbreak Detection for Large Language Models using Token-Level Logits

DGX agent

arXiv:2604.01473v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are powerful tools for answering user queries, yet they remain highly vulnerable to jailbreak attacks. Existing g

model-releasesarxiv-cs-ai
17 Apr 2026
Model Releases

Sharing my current setup to run Qwen3.6 locally in a good agentic setup (Pi + llama.cpp). Should give you a good overview of how good local …

DGX agent

Sharing my current setup to run Qwen3.6 locally in a good agentic setup (Pi + llama.cpp). Should give you a good overview of how good local agents are today: # Start llama.cpp server: llama-server -hf

model-releasesclem-delangue--x
17 Apr 2026
Model Releases

Shuffle the Context: RoPE-Perturbed Self-Distillation for Long-Context Adaptation

DGX agent

arXiv:2604.14339v1 Announce Type: new Abstract: Large language models (LLMs) increasingly operate in settings that require reliable long-context understanding, such as retrieval-augmented generation a

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

SOLIS: Physics-Informed Learning of Interpretable Neural Surrogates for Nonlinear Systems

DGX agent

arXiv:2604.14879v1 Announce Type: new Abstract: Nonlinear system identification must balance physical interpretability with model flexibility. Classical methods yield structured, control-relevant mode

model-releasesarxiv-cs-lg
17 Apr 2026
Model Releases

SpaceX completed two Starlink launches about 19 hours apart, pushing the number of Starlink satellites launched in 2026 past 1,000. Starlink…

DGX agent

SpaceX completed two Starlink launches about 19 hours apart, pushing the number of Starlink satellites launched in 2026 past 1,000. Starlink 10-24 from Florida put 29 V2 Mini satellites into orbit wit

model-releaseselon-musk--x
17 Apr 2026
Model Releases

Speak, Segment, Track, Navigate: An Interactive System for Video-Guided Skull-Base Surgery

DGX agent

arXiv:2603.16024v2 Announce Type: replace Abstract: We introduce a speech-guided embodied agent framework for video-guided skull base surgery that dynamically executes perception and image-guidance ta

model-releasesarxiv-cs-cv
17 Apr 2026
← Previous
1…422423424425426…465
Next →