AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,553 results
20 Jul 2026

IssueBench is our internal benchmark for evaluating Engine (a continual learning agent in LangSmith) This blog by @nick_bray dives into why …

Model ReleasesDGX agent

**IssueBench** is an internal benchmark created by LangSmith to assess the performance of *Engine*, a continual‑learning agent that scans other agents’ traces to identify, cluster, and fix issues. In

Open ecosystem or walled garden? For @anthropicai's @katelyn_lesse and @angjiang the answer is clear. 'We actually aren't precious about “Yo…

Model ReleasesDGX agent

Open ecosystem or walled garden? For @anthropicai's @katelyn_lesse and @angjiang the answer is clear. 'We actually aren't precious about “You should run these things on our infrastructure.' In practic

People are wasting their AI subscriptions. They don’t realize how powerful AI can be in achieving their own goals. They’re asking AI to repl…


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

People are wasting their AI subscriptions. They don’t realize how powerful AI can be in achieving their own goals. They’re asking AI to reply to emails with zero goals in mind and nothing laddering up

Was very cool to hear about the reasons people love Sol. We're doing the promotion again, except this time for ChatGPT Work: Tweet what you …

Model ReleasesDGX agent

Was very cool to hear about the reasons people love Sol. We're doing the promotion again, except this time for ChatGPT Work: Tweet what you love about ChatGPT Work, claim 100 in free credits, get more

Who’s Afraid of Chinese Models?

Model ReleasesDGX agent

Who’s Afraid of Chinese Models? Interesting proposal from Ben Thompson that both addresses the hypocrisy of labs outlawing distillation against their models despite training on unlicensed data, and co

16 Jul 2026

2D Rotary Position Embedding for Scene Text Recognition with Transformers

Model ReleasesDGX agent

arXiv:2607.13458v1 Announce Type: new Abstract: Scene Text Recognition (STR) remains challenging due to the diversity of text appearances, including curvature, rotation, and perspective distortion. Re

A Comparative Evaluation of Large Vision-Language Models for 2D Object Detection under SOTIF Conditions

Model ReleasesDGX agent

arXiv:2601.22830v2 Announce Type: replace Abstract: Reliable environmental perception remains one of the main obstacles for safe operation of automated vehicles. Safety of the Intended Functionality (

A plug-and-play approach with fast uncertainty quantification for weak lensing mass mapping

Model ReleasesDGX agent

arXiv:2603.22006v2 Announce Type: replace-cross Abstract: Upcoming stage-IV surveys such as Euclid and Rubin will deliver vast amounts of high-precision data, opening new opportunities to constrain co

A Self-Evolving Agent for Longitudinal Personal Health Management

Model ReleasesDGX agent

arXiv:2607.13940v1 Announce Type: new Abstract: Personal health management unfolds over repeated encounters, yet most health AI systems treat each request in isolation. We developed HealthClaw, an ope

Accuracy Without Grounding: Diagnosing Visual Dependency Dissociation in Video LLM Benchmarks

Model ReleasesDGX agent

arXiv:2607.13305v1 Announce Type: cross Abstract: Benchmark accuracy in video large language models (LLMs) is often treated as evidence of visual understanding. We audit this assumption across twenty

Advancing Multimodal Judge Models through a Capability-Oriented Benchmark and MCTS-Driven Data Generation

Model ReleasesDGX agent

arXiv:2603.00546v2 Announce Type: replace Abstract: Using Multimodal Large Language Models (MLLMs) as judges to achieve precise and consistent evaluations has gradually become an emerging paradigm acr

AgentCompass: A Unified Evaluation Infrastructure for Agent Capabilities

Model ReleasesDGX agent

arXiv:2607.13705v1 Announce Type: new Abstract: As Large Language Models (LLMs) evolve into autonomous agents, the need for unified evaluation infrastructure becomes critical. However, current evaluat

Agora: Collective and Permissionless Internet-Scale Pretraining of Large Language Models

Model ReleasesDGX agent

arXiv:2607.13332v1 Announce Type: new Abstract: Training large language models at the multi-billion to trillion parameter scale is confined to datacenters, where data-parallel (DP) and model-parallel

AIMO Interpretability Challenge

Model ReleasesDGX agent

arXiv:2607.13899v1 Announce Type: new Abstract: We propose the AIMO Interpretability Challenge, a competition on distinguishing robust from spurious reasoning in frontier mathematical language models

An Explainable Agentic System for Detection of Conversational Scams with Summary-Based Memory

Model ReleasesDGX agent

arXiv:2607.11707v2 Announce Type: replace-cross Abstract: Following the rapid progress of generative Artificial Intelligence, there is a growing threat posed by conversational scams. These scams often

Analogical Deep Research: Retrieving and Integrating Historical Analogies for Foresight Analysis

Model ReleasesDGX agent

arXiv:2607.13602v1 Announce Type: cross Abstract: Systematic comparisons between current situations and structurally similar past events in the historical, i.e., historical analogies, is among the mos

AnomExpert: Identifying and Selecting Anatomical Planes for Prenatal Ultrasound Anomaly Diagnosis

Model ReleasesDGX agent

arXiv:2607.13409v1 Announce Type: new Abstract: Life-limiting congenital anomalies require accurate prenatal diagnosis for appropriate clinical decision-making. Prenatal ultrasound (US) examinations i

Approximation of solutions of parameter-dependent problems by residual neural networks

Model ReleasesDGX agent

arXiv:2607.13574v1 Announce Type: cross Abstract: We develop a convergent scheme to train neural networks involving analytic activation functions based on gradient flows. Convergence properties are gu

Ask Before You Diagnose: Safe-Psych, a Sequential Evaluation Benchmark for LLMs in Psychiatry

Model ReleasesDGX agent

arXiv:2607.13036v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used for decision support in healthcare, but clinical evidence is often incomplete or evolving. When the

Automatic Differentiation from Scratch: How PyTorch Computes Gradients in Physics-Informed Neural Networks

Model ReleasesDGX agent

arXiv:2607.13042v1 Announce Type: new Abstract: This paper traces, with explicit numerical values, how PyTorch's automatic differentiation (AD) engine computes gradients for Physics-Informed Neural Ne

b10036

Model ReleasesDGX agent

opencl: disable FA and MoE weights repack to work around compiler issues for Adreno 850 GPU (#25745) opencl: workaround for A850 compiler compat opencl: fix DX compiler version parsing and cleanup Co-

b10038

Model ReleasesDGX agent

ci : add official website link to release notes (#25728) Assisted-by: pi:llama.cpp/Qwen3.6-27B Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI en

Barnamala: Parameter-Efficient Handwritten Devanagari Recognition at Benchmark Saturation

Model ReleasesDGX agent

arXiv:2607.13689v1 Announce Type: cross Abstract: We built a compact convolutional network (1.11 M parameters) for 46-class DHCD Devanagari recognition and reached 99.73%, the highest reported at 15.6

Baselines Before Architecture: Evaluating Coding Agents for Autonomous Penetration Testing

Model ReleasesDGX agent

arXiv:2607.13085v1 Announce Type: cross Abstract: Recent autonomous penetration testing papers report high benchmark scores while adding multi-component security harnesses around frontier LLMs. Becaus

BenthiCat: An opti-acoustic dataset for advancing benthic classification and habitat mapping

Model ReleasesDGX agent

arXiv:2510.04876v3 Announce Type: replace Abstract: Benthic habitat mapping is fundamental for understanding marine ecosystems, guiding conservation efforts, and supporting sustainable resource manage

Beyond Description: Cognitively Benchmarking Fine-Grained Action for Embodied Agents

Model ReleasesDGX agent

arXiv:2511.18685v4 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) show promising results as decision-making engines for embodied agents operating in complex, physical enviro

Calibrated Closed-Form Uncertainty for Radiative Gaussian Splatting in Sparse-View CT

Model ReleasesDGX agent

arXiv:2607.13682v1 Announce Type: new Abstract: Radiative Gaussian splatting has made sparse-view CT reconstruction fast, but existing methods output point estimates with no notion of where the recons

Can We Steer the Black-Box? Towards Controllability-Centric Evaluation of Recommender Systems with Collaborative Agents

Model ReleasesDGX agent

arXiv:2607.13418v1 Announce Type: cross Abstract: Recommender systems operate as Black-Boxes, leaving users and regulators unable to steer their outputs toward specific intentions or audit their behav

CASA-SDF: Curriculum-Aware Spatial Adaptation with Curvature-Guided Density for Neural Implicit Surface Reconstruction

Model ReleasesDGX agent

arXiv:2607.13492v1 Announce Type: new Abstract: Neural implicit representations have emerged as a powerful paradigm for 3D reconstruction. However, high-fidelity indoor surface reconstruction remains

CAVA: Canonical Action Verification and Attestation for Runtime Governance of Agentic AI Systems

Model ReleasesDGX agent

arXiv:2607.13716v1 Announce Type: new Abstract: Agentic AI systems increasingly act through heterogeneous runtimes: local coding hooks, SDK tools, browser automation, managed-agent traces, API gateway

CoDiffGRN: Rethinking Gene Regulatory Network Inference via the BEELINE-KGC Benchmark and Co-evolutionary Discrete Diffusion

Model ReleasesDGX agent

arXiv:2607.13120v1 Announce Type: cross Abstract: Inferring gene regulatory networks (GRNs) from single-cell transcriptomic data is crucial for biological discovery, yet existing approaches suffer fro

COLMAR: Cooperative View Policy Learning for Multi-Agent Active 3D Reconstruction

Model ReleasesDGX agent

arXiv:2607.13524v1 Announce Type: new Abstract: Active 3D reconstruction requires selecting informative viewpoints under limited sensing budgets. In multi-agent settings, coordination inefficiencies s

Compaction as Epistemic Failure: How Agentic LLM Tools Fabricate Confirmed Results from Killed Processes

Model ReleasesDGX agent

arXiv:2607.13071v1 Announce Type: cross Abstract: Agentic LLM coding tools compress long session histories into compaction summaries that subsequent sessions inherit as ground truth. This paper docume

Constraint-Driven Model Optimization: An Industry Framework for Selecting Compression and Acceleration Techniques in Modern Machine Learning Systems

Model ReleasesDGX agent

arXiv:2607.13735v1 Announce Type: new Abstract: The rapid deployment of machine learning systems across cloud, edge, and enterprise environments has brought model optimization to the forefront of syst

Continuously Evolving Deepfake Detection: An Architecture and Public-Benchmark Evaluation of a Dynamic Detection System

Model ReleasesDGX agent

arXiv:2607.13234v1 Announce Type: cross Abstract: Deepfake detectors that achieve near-perfect scores on academic benchmarks collapse on real-world content: recent in-the-wild evaluations report AUC d

cuda: extract Q1_0 elements via __byte_perm by dfriehs · Pull Request #25628 · ggml-org/llama.cpp

Model ReleasesDGX agent

I don't have the ability to access Reddit posts or browse specific URLs. To provide you with an accurate factual summary for your knowledge base, I would need either: 1. The actual content/text from t

DarwinLM: Evolutionary Structured Pruning of Large Language Models

Model ReleasesDGX agent

arXiv:2502.07780v4 Announce Type: replace Abstract: Large Language Models (LLMs) have achieved significant success across various NLP tasks. However, their massive computational costs limit their wide

Data-Efficient Adaptation of LLMs via Attention Head Reweighting

Model ReleasesDGX agent

arXiv:2607.13425v1 Announce Type: cross Abstract: Learning effectively from limited data is critical in domains like security where labeled examples are scarce. Large language models (LLMs) have demon

DeepLoop: Depth Scaling for Looped Transformers

Model ReleasesDGX agent

arXiv:2607.13491v1 Announce Type: cross Abstract: Looped Transformers scale sequential computation by applying a compact stack of physical blocks for multiple rounds, increasing unrolled depth without

Design, Modeling and Experimental Validation of a Miniature Hybrid Underwater Glider With Large-Range Foldable Deflectable Wings

Model ReleasesDGX agent

arXiv:2607.13622v1 Announce Type: new Abstract: Miniature hybrid underwater gliders have attracted increasing attention for long-endurance ocean observation and confined-space inspection. Large-range

Detector Confidence Signals Presence Rather Than Occlusion in Cluttered Manipulation

Model ReleasesDGX agent

arXiv:2607.13361v1 Announce Type: new Abstract: Occlude a named object until about an eighth of it remains visible, and an open-vocabulary detector's confidence that the object is present barely chang

DevicesWorld: Benchmarking Cross-Device Agents in Heterogeneous Environments

Model ReleasesDGX agent

arXiv:2607.13465v1 Announce Type: cross Abstract: LLM-based agents have rapidly improved at operating individual digital environments such as mobile applications, desktop systems, and smart homes. How

Discriminative Barrier Functions for Safe Adversarial Imitation Learning from Observation

Model ReleasesDGX agent

arXiv:2607.13938v1 Announce Type: new Abstract: Inverse Reinforcement Learning (IRL) algorithms are powerful tools for learning from and generalizing expert demonstrations, but they often rely on unco

DNA: Dual-stage Native Attribution for Generated Image Source Tracing

Model ReleasesDGX agent

arXiv:2607.13685v1 Announce Type: new Abstract: The rapid evolution of image generation has produced numerous within-family variants, making source-model attribution of suspect images increasingly imp

Do Agent Optimizers Compound? A Continual-Learning Evaluation on Terminal-Bench 2.0

Model ReleasesDGX agent

arXiv:2607.14004v1 Announce Type: new Abstract: Most reported gains from agent-optimization methods are one-shot: an agent is optimized against a fixed benchmark and the resulting improvement is repor

Dynamical Vehicle Orienteering Problem for Multi-Rotor Unmanned Aerial Vehicles

Model ReleasesDGX agent

arXiv:2607.13789v1 Announce Type: new Abstract: This paper introduces the Dynamical Vehicle Orienteering Problem (DVOP), a generalization of the Orienteering Problem (OP). The OP maximizes the reward

Efficient Text-to-Audio Generation via Pruning

Model ReleasesDGX agent

arXiv:2607.13330v1 Announce Type: cross Abstract: Diffusion-based text-to-audio generative models such as AudioLDM achieve high perceptual quality and strong semantic consistency; however, their pract

EgoHTR: Egocentric 4D Demonstrations of Human Terrain Traversal

Model ReleasesDGX agent

arXiv:2607.13472v1 Announce Type: cross Abstract: Deploying humanoid robots in unstructured terrain remains an open problem. While classic reinforcement learning struggles with the sheer complexity of

EgoProceVQA: A Novel Egocentric Procedural Understanding Task with Self-Skill-Exploration Agent

Model ReleasesDGX agent

arXiv:2607.13792v1 Announce Type: new Abstract: Most daily activities are inherently procedural. However, existing evaluations for egocentric video understanding seldom address procedural understandin

Evaluating Frontier AI Agents as Autonomous Clinical Security Auditors

Model ReleasesDGX agent

arXiv:2607.13411v1 Announce Type: cross Abstract: Clinical AI models can expose patients to harm when adversarial vulnerabilities go undetected, yet formal security auditing requires statistical exper

Evaluating Vision Foundation Models for Pixel and Object Classification in Microscopy

Model ReleasesDGX agent

arXiv:2603.19802v2 Announce Type: replace Abstract: Deep learning underlies most modern approaches and tools in computer vision, including biomedical imaging. However, for interactive semantic segment

Exploratory, Communicative, and Deployable: Vision-Driven Embodied Agents for Open-World Mobile Manipulation

Model ReleasesDGX agent

arXiv:2607.13653v1 Announce Type: new Abstract: Real-world deployment of embodied agents requires active exploration, visual grounding, and interactive intent disambiguation. However, existing framewo

EXPLORE: Exploration with Guided Search for Analog Topology Generation using Language Models

Model ReleasesDGX agent

arXiv:2607.13416v1 Announce Type: new Abstract: Automating analog circuit topology design is essential to reduce the extensive manual effort required to meet increasingly diverse and customized applic

ExTernD: Expanded-Rank Ternary Decomposition Ternary LLM PTQ with Accuracy Approaching Any Quantization Level

Model ReleasesDGX agent

arXiv:2607.13511v1 Announce Type: cross Abstract: We introduce ExTernD (Expanded-rank Ternary Decomposition), a post-training factorization of each LLM weight matrix A in R^{m imes n} into A approx B

FastCentNN: Accelerating Centroid Neural Network with Entropy Proxy

Model ReleasesDGX agent

arXiv:2607.13613v1 Announce Type: cross Abstract: Centroid neural network (CentNN) is an unsupervised competitive learning algorithm in which centroid splitting is triggered only after strict local st

Fine-grained CLIP fine-tuning with self-annotated region alignment

Model ReleasesDGX agent

arXiv:2607.13661v1 Announce Type: new Abstract: Contrastive Language-Image Pre-training (CLIP) has been shown to have limitations in its fine-grained dense feature representation, due to its pre-train

Fire as a Service: Augmenting Robot Simulators with Thermally and Visually Accurate Fire Dynamics

Model ReleasesDGX agent

arXiv:2603.19063v2 Announce Type: replace Abstract: Most existing robot simulators prioritize rigid-body dynamics and photorealistic rendering, but largely neglect the thermally and optically complex

FM^2: Unified Federated Foundation Models for Heterogeneous Multimodal Medical Imaging

Model ReleasesDGX agent

arXiv:2607.13386v1 Announce Type: new Abstract: Building foundation models for medical imaging requires pooling data across institutions, yet privacy regulations prohibit centralized aggregation. Exis

From Language to Navigation Goals: A Vision-Language Approach for Semantic Navigation of Mobile Robots Using RGB-D Perception

Model ReleasesDGX agent

arXiv:2607.13624v1 Announce Type: cross Abstract: Natural language interaction provides an intuitive way for non-expert users to communicate with robotic platforms. However, transforming user requests

From Prediction to Collaboration: Interactive Symbolic Music Analysis

Model ReleasesDGX agent

arXiv:2607.13587v1 Announce Type: cross Abstract: Automatic symbolic music analysis has made substantial progress, yet existing systems are typically designed for a single mode of use, such as full-sc

← Previous
1…7980818283…376
Next →