AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

88,271Total entries
1Added by human
88,270Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,519 results
17 Apr 2026

IG-Search: Step-Level Information Gain Rewards for Search-Augmented Reasoning

SafetyDGX agent

arXiv:2604.15148v1 Announce Type: cross Abstract: Reinforcement learning has emerged as an effective paradigm for training large language models to perform search-augmented reasoning. However, existin

Learning Adaptive Reasoning Paths for Efficient Visual Reasoning

SafetyDGX agent

arXiv:2604.14568v1 Announce Type: cross Abstract: Visual reasoning models (VRMs) have recently shown strong cross-modal reasoning capabilities by integrating visual perception with language reasoning.

Leveraging graph neural networks and mobility data for COVID-19 forecasting

ResearchDGX agent

arXiv:2501.11711v2 Announce Type: replace Abstract: The COVID-19 pandemic has claimed millions of lives, spurring the development of diverse forecasting models. In this context, the true utility of co

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

LLM Predictive Scoring and Validation: Inferring Experience Ratings from Unstructured Text

Model ReleasesDGX agent

arXiv:2604.14321v1 Announce Type: new Abstract: We tasked GPT-4.1 to read what baseball fans wrote about their game-day experience and predict the overall experience rating each fan gave on a 0-10 sur

MambaSL: Exploring Single-Layer Mamba for Time Series Classification

ResearchDGX agent

arXiv:2604.15174v1 Announce Type: new Abstract: Despite recent advances in state space models (SSMs) such as Mamba across various sequence domains, research on their standalone capacity for time serie

Maximal Brain Damage Without Data or Optimization: Disrupting Neural Networks via Sign-Bit Flips

Model ReleasesDGX agent

arXiv:2502.07408v2 Announce Type: replace-cross Abstract: Deep Neural Networks (DNNs) can be catastrophically disrupted by flipping only a handful of parameter bits. We introduce Deep Neural Lesion (D

ModuSeg: Decoupling Object Discovery and Semantic Retrieval for Training-Free Weakly Supervised Segmentation

Model ReleasesDGX agent

arXiv:2604.07021v2 Announce Type: replace Abstract: Weakly supervised semantic segmentation aims to achieve pixel-level predictions using image-level labels. Existing methods typically entangle semant

On the Expressive Power and Limitations of Multi-Layer SSMs

ResearchDGX agent

arXiv:2604.14501v1 Announce Type: new Abstract: We study the expressive power and limitations of multi-layer state-space models (SSMs). First, we show that multi-layer SSMs face fundamental limitation

Open-Set Vein Biometric Recognition with Deep Metric Learning

Model ReleasesDGX agent

arXiv:2604.14874v1 Announce Type: new Abstract: Most state-of-the-art vein recognition methods rely on closed-set classification, which inherently limits their scalability and prevents the adaptive en

Preconditioned Test-Time Adaptation for Out-of-Distribution Debiasing in Narrative Generation

SafetyDGX agent

arXiv:2603.13683v2 Announce Type: replace Abstract: Although debiased large language models (LLMs) excel at handling known or low-bias prompts, they often fail on unfamiliar and high-bias prompts. We

Query pipeline optimization for cancer patient question answering systems

Model ReleasesDGX agent

arXiv:2412.14751v2 Announce Type: replace Abstract: Retrieval-augmented generation (RAG) mitigates hallucination in Large Language Models (LLMs) by using query pipelines to retrieve relevant external

ReviewGrounder: Improving Review Substantiveness with Rubric-Guided, Tool-Integrated Agents

Model ReleasesDGX agent

arXiv:2604.14261v1 Announce Type: new Abstract: The rapid rise in AI conference submissions has driven increasing exploration of large language models (LLMs) for peer review support. However, LLM-base

Reward-Aware Trajectory Shaping for Few-step Visual Generation

SafetyDGX agent

arXiv:2604.14910v1 Announce Type: new Abstract: Achieving high-fidelity generation in extremely few sampling steps has long been a central goal of generative modeling. Existing approaches largely rely

SafeHarness: Lifecycle-Integrated Security Architecture for LLM-based Agent Deployment

Model ReleasesDGX agent

arXiv:2604.13630v1 Announce Type: cross Abstract: The performance of large language model (LLM) agents depends critically on the execution harness, the system layer that orchestrates tool use, context

Sharing my current setup to run Qwen3.6 locally in a good agentic setup (Pi + llama.cpp). Should give you a good overview of how good local …

Model ReleasesDGX agent

Sharing my current setup to run Qwen3.6 locally in a good agentic setup (Pi + llama.cpp). Should give you a good overview of how good local agents are today: # Start llama.cpp server: llama-server -hf

Tesla is officially launching in Estonia. The company is holding an opening event on April 24th at Ülemiste Center 'Bring your friends and f…

IndustryDGX agent

Tesla is officially launching in Estonia. The company is holding an opening event on April 24th at Ülemiste Center 'Bring your friends and family to see our models, take part in the day’s activities a

Text2Arch: A Dataset for Generating Scientific Architecture Diagrams from Natural Language Descriptions

ApplicationsDGX agent

arXiv:2604.14941v1 Announce Type: new Abstract: Communicating complex system designs or scientific processes through text alone is inefficient and prone to ambiguity. A system that automatically gener

Theory of Mind in Action: The Instruction Inference Task in Dynamic Human-Agent Collaboration

Model ReleasesDGX agent

arXiv:2507.02935v2 Announce Type: replace Abstract: Successful human-agent teaming relies on an agent being able to understand instructions given by a (human) principal. In many cases, an instruction

Towards AI-assisted Neutrino Flavor Theory Design

AgentsDGX agent

arXiv:2506.08080v2 Announce Type: replace-cross Abstract: Particle physics theories, such as those which explain neutrino flavor mixing, arise from a vast landscape of model-building possibilities. A

Towards Deploying VLA without Fine-Tuning: Plug-and-Play Inference-Time VLA Policy Steering via Embodied Evolutionary Diffusion

SafetyDGX agent

arXiv:2511.14178v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models have demonstrated significant potential in real-world robotic manipulation. However, pre-trained VLA policies st

Uncovering the Fragility of Trustworthy LLMs through Chinese Textual Ambiguity

Model ReleasesDGX agent

arXiv:2507.23121v2 Announce Type: replace Abstract: In this work, we study a critical research problem regarding the trustworthiness of large language models (LLMs): how LLMs behave when encountering

VeriGraphi: A Multi-Agent Framework of Hierarchical RTL Generation for Large Hardware Designs

Model ReleasesDGX agent

arXiv:2604.14550v1 Announce Type: cross Abstract: Generating synthesizable Verilog for large, hierarchical hardware designs remains a significant challenge for large language models (LLMs), which stru

When Fairness Metrics Disagree: Evaluating the Reliability of Demographic Fairness Assessment in Machine Learning

SafetyDGX agent

arXiv:2604.15038v1 Announce Type: cross Abstract: The evaluation of fairness in machine learning systems has become a central concern in high-stakes applications, including biometric recognition, heal

When PCOS Meets Eating Disorders: An Explainable AI Approach to Detecting the Hidden Triple Burden

Model ReleasesDGX agent

arXiv:2604.14356v1 Announce Type: new Abstract: Women with polycystic ovary syndrome (PCOS) face substantially elevated risks of body image distress, disordered eating, and metabolic challenges, yet e

yeah, what they said 🤝 a decent harness gets you an actually functioning agent now + teams that actually invest time in their harness+probl…

AgentsDGX agent

yeah, what they said 🤝 a decent harness gets you an actually functioning agent now + teams that actually invest time in their harness+problem design, choosing good infra, self-improvement loops, data

16 Apr 2026

AeTHERON: Autoregressive Topology-aware Heterogeneous Graph Operator Network for Fluid-Structure Interaction

Model ReleasesDGX agent

arXiv:2604.13369v1 Announce Type: cross Abstract: Surrogate modeling of body-driven fluid flows where immersed moving boundaries couple structural dynamics to chaotic, unsteady fluid phenomena remains

AI Powered Image Analysis for Phishing Detection

Model ReleasesDGX agent

arXiv:2604.13555v1 Announce Type: new Abstract: Phishing websites now rely heavily on visual imitation-copied logos, similar layouts, and matching colours-to avoid detection by text- and URL-based sys

Analog Optical Inference on Million-Record Mortgage Data

Model ReleasesDGX agent

arXiv:2604.13251v1 Announce Type: new Abstract: Analog optical computers promise large efficiency gains for machine learning inference, yet no demonstration has moved beyond small-scale image benchmar

AudioX: A Unified Framework for Anything-to-Audio Generation

Model ReleasesDGX agent

arXiv:2503.10522v4 Announce Type: replace-cross Abstract: Audio and music generation based on flexible multimodal control signals is a widely applicable topic, with the following key challenges: 1) a

CLIP Architecture for Abdominal CT Image-Text Alignment and Zero-Shot Learning: Investigating Batch Composition and Data Scaling

SafetyDGX agent

arXiv:2604.13561v1 Announce Type: new Abstract: Vision-language models trained with contrastive learning on paired medical images and reports show strong zero-shot diagnostic capabilities, yet the eff

Coherence in the brain unfolds across separable temporal regimes

Model ReleasesDGX agent

arXiv:2512.20481v4 Announce Type: replace-cross Abstract: To maintain coherence in language, the brain must satisfy key competing temporal demands: the gradual accumulation of meaning across extended

Concrete Jungle: Towards Concreteness Paved Contrastive Negative Mining for Compositional Understanding

ResearchDGX agent

arXiv:2604.13313v1 Announce Type: new Abstract: Vision-Language Models demonstrate remarkable capabilities but often struggle with compositional reasoning, exhibiting vulnerabilities regarding word or

Context Sensitivity Improves Human-Machine Visual Alignment

SafetyDGX agent

arXiv:2604.13883v1 Announce Type: new Abstract: Modern machine learning models typically represent inputs as fixed points in a high-dimensional embedding space. While this approach has been proven pow

Correct Chains, Wrong Answers: Dissociating Reasoning from Output in LLM Logic

Model ReleasesDGX agent

arXiv:2604.13065v1 Announce Type: new Abstract: LLMs can execute every step of chain-of-thought reasoning correctly and still produce wrong final answers. We introduce the Novel Operator Test, a bench

DeEscalWild: A Real-World Benchmark for Automated De-Escalation Training with SLMs

Model ReleasesDGX agent

arXiv:2604.13075v1 Announce Type: new Abstract: Effective de-escalation is critical for law enforcement safety and community trust, yet traditional training methods lack scalability and realism. While

DiffMagicFace: Identity Consistent Facial Editing of Real Videos

ResearchDGX agent

arXiv:2604.13841v1 Announce Type: new Abstract: Text-conditioned image editing has greatly benefitted from the advancements in Image Diffusion Models. However, extending these techniques to facial vid

DiT as Real-Time Rerenderer: Streaming Video Stylization with Autoregressive Diffusion Transformer

ResearchDGX agent

arXiv:2604.13509v1 Announce Type: new Abstract: Recent advances in video generation models has significantly accelerated video generation and related downstream tasks. Among these, video stylization h

Enhancing Mixture-of-Experts Specialization via Cluster-Aware Upcycling

ResearchDGX agent

arXiv:2604.13508v1 Announce Type: new Abstract: Sparse Upcycling provides an efficient way to initialize a Mixture-of-Experts (MoE) model from pretrained dense weights instead of training from scratch

Evaluating LLM-Based Translation of a Low-Resource Technical Language: The Medical and Philosophical Greek of Galen

Model ReleasesDGX agent

arXiv:2602.24119v2 Announce Type: replace Abstract: Purpose: This study evaluates the quality of commercial large language model (LLM) machine translation (MT) for Ancient Greek technical prose and be

EVE: A Domain-Specific LLM Framework for Earth Intelligence

Model ReleasesDGX agent

arXiv:2604.13071v1 Announce Type: new Abstract: We introduce Earth Virtual Expert (EVE), the first open-source, end-to-end initiative for developing and deploying domain-specialized LLMs for Earth Int

Gemini can now pull from Google Photos to generate personalized images

Model ReleasesDGX agent

Google's Personal Intelligence feature, which lets Gemini pull data from apps like Google Photos to offer responses tailored to you, can now use that data and its Nano Banana 2 image model to create i

HiVLA: A Visual-Grounded-Centric Hierarchical Embodied Manipulation System

ApplicationsDGX agent

arXiv:2604.14125v1 Announce Type: new Abstract: While end-to-end Vision-Language-Action (VLA) models offer a promising paradigm for robotic manipulation, fine-tuning them on narrow control data often

I have found that asking for a sestina regularly triggers Opus 4.7's safety guardrails. The forbidden poetic form!

SafetyDGX agent

Ethan Mollick reported that requesting Claude Opus 4.7 to write sestinas—a complex poetic form with strict structural requirements—frequently triggers the model's safety guardrails, suggesting the AI

MAny: Merge Anything for Multimodal Continual Instruction Tuning

Model ReleasesDGX agent

arXiv:2604.14016v1 Announce Type: new Abstract: Multimodal Continual Instruction Tuning (MCIT) is essential for sequential task adaptation of Multimodal Large Language Models (MLLMs) but is severely r

Med-CAM: Minimal Evidence for Explaining Medical Decision Making

SafetyDGX agent

arXiv:2604.13695v1 Announce Type: new Abstract: Reliable and interpretable decision-making is essential in medical imaging, where diagnostic outcomes directly influence patient care. Despite advances

MERRIN: A Benchmark for Multimodal Evidence Retrieval and Reasoning in Noisy Web Environments

Model ReleasesDGX agent

arXiv:2604.13418v1 Announce Type: new Abstract: Motivated by the underspecified, multi-hop nature of search queries and the multimodal, heterogeneous, and often conflicting nature of real-world web re

More on my blog, including results from the previously secret 'flamingo on a unicycle' test https://simonwillison.net/2026/Apr/16/qwen-beats…

Model ReleasesDGX agent

Simon Willison discusses results from a 'flamingo on a unicycle' test on his blog, likely comparing AI model performance including Qwen. The post appears to reference previously undisclosed or unconve

One Token per Highly Selective Frame: Towards Extreme Compression for Long Video Understanding

Local AiDGX agent

arXiv:2604.14149v1 Announce Type: new Abstract: Long video understanding is inherently challenging for vision-language models (VLMs) because of the extensive number of frames. With each video frame ty

OneHOI: Unifying Human-Object Interaction Generation and Editing

ResearchDGX agent

arXiv:2604.14062v1 Announce Type: new Abstract: Human-Object Interaction (HOI) modelling captures how humans act upon and relate to objects, typically expressed as triplets. Existing approaches split

OPTED: Open Preprocessed Trachoma Eye Dataset Using Zero-Shot SAM 3 Segmentation

Model ReleasesDGX agent

arXiv:2603.06885v2 Announce Type: replace Abstract: Trachoma remains the leading infectious cause of blindness worldwide, with Sub-Saharan Africa bearing over 85% of the global burden and Ethiopia alo

Opus 4.7 is now supported in Hermes Agent 🚀🚀

Model ReleasesDGX agent

Opus 4.7 is now supported in Hermes Agent 🚀🚀 Introducing Claude Opus 4.7, our most capable Opus model yet. It handles long-running tasks with more rigor, follows instructions more precisely, and verif

Out of Context: Reliability in Multimodal Anomaly Detection Requires Contextual Inference

Model ReleasesDGX agent

arXiv:2604.13252v1 Announce Type: new Abstract: Anomaly detection aims to identify observations that deviate from expected behavior. Because anomalous events are inherently sparse, most frameworks are

Red Skills or Blue Skills? A Dive Into Skills Published on ClawHub

Model ReleasesDGX agent

arXiv:2604.13064v1 Announce Type: new Abstract: Skill ecosystems have emerged as an increasingly important layer in Large Language Model (LLM) agent systems, enabling reusable task packaging, public d

Replit Agent 4 is even smarter now with Claude Opus 4.7! 50% off for a limited time. Go try it now ↓

Model ReleasesDGX agent

Replit Agent 4 has been upgraded to use Claude Opus 4.7, an advanced AI model, enhancing its code generation and problem-solving capabilities. The company is offering a 50% discount for a limited time

Response.

Model ReleasesDGX agent

Response. Hey Ethan! Sean here, PM on http://Claude.ai - thanks for the feedback. This isn't a router, this is the model being trained to decide when to think based on the context -- we've been runnin

Rhetorical Questions in LLM Representations: A Linear Probing Study

Local AiDGX agent

arXiv:2604.14128v1 Announce Type: new Abstract: Rhetorical questions are asked not to seek information but to persuade or signal stance. How large language models internally represent them remains unc

Robust Ultra Low-Bit Post-Training Quantization via Stable Diagonal Curvature Estimate

ResearchDGX agent

arXiv:2604.13806v1 Announce Type: new Abstract: Large Language Models (LLMs) are widely used across many domains, but their scale makes deployment challenging. Post-Training Quantization (PTQ) reduces

RPS: Information Elicitation with Reinforcement Prompt Selection

Model ReleasesDGX agent

arXiv:2604.13817v1 Announce Type: new Abstract: Large language models (LLMs) have shown remarkable capabilities in dialogue generation and reasoning, yet their effectiveness in eliciting user-known bu

SparseBalance: Load-Balanced Long Context Training with Dynamic Sparse Attention

Model ReleasesDGX agent

arXiv:2604.13847v1 Announce Type: new Abstract: While sparse attention mitigates the computational bottleneck of long-context LLM training, its distributed training process exhibits extreme heterogene

Stability Principle Underlying Passive Dynamic Walking of Rimless Wheel

ResearchDGX agent

arXiv:2604.13530v1 Announce Type: new Abstract: Rimless wheels are known as the simplest model for passive dynamic walking. It is known that the passive gait generated only by gravity effect always be

← Previous
1…429430431432433…1059
Next →