AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,569 results
2 Jul 2026

Can Agents Generalize to the Open World? Unveiling the Fragility of Static Training in Tool Use

Model ReleasesDGX agent

arXiv:2607.01084v1 Announce Type: new Abstract: While Large Language Model (LLM) agents demonstrate proficiency in static benchmarks, their deployment in real-world scenarios is hindered by the dynami

Clinician-Level Agreement Without Clinical Caution: LLM Evaluator Limits in Medical AI Benchmarking

Model ReleasesDGX agent

arXiv:2607.01103v1 Announce Type: new Abstract: Open-response evaluation provides stronger clinical validity than multiple-choice benchmarks but creates a scoring bottleneck that motivates automated L

Comparing Large Language Models on Scrum Certification-Style Questions: Accuracy, Stability, and Error Patterns


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

arXiv:2607.00048v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in exam- and certification-style question answering tasks, where their ability to retrieve, interpr

Computer vision-based neural networks for radioisotope identification in urban environments

Model ReleasesDGX agent

arXiv:2607.00270v1 Announce Type: cross Abstract: Algorithm development for radioisotope identification in mobile urban search scenarios face significant challenges from non-uniform backgrounds, momen

Continual learning is probably the biggest barrier to explosive AI adoption (& may have big implications for recursive self-improvement as w…

Model ReleasesDGX agent

Continual learning is probably the biggest barrier to explosive AI adoption (& may have big implications for recursive self-improvement as well) As long as you deal with amnesiac models that require h

CoT-X: An Adaptive Framework for Cross-Model Chain-of-Thought Transfer and Optimization

Model ReleasesDGX agent

arXiv:2511.05747v3 Announce Type: replace Abstract: Chain-of-Thought (CoT) reasoning enhances the problem-solving ability of large language models (LLMs) but leads to substantial inference overhead, l

CPDDNet: Color-Polarization Denoising and Demosaicking Network

Model ReleasesDGX agent

arXiv:2607.01100v1 Announce Type: new Abstract: Color-polarization imaging using a color-polarization filter array (CPFA) sensor captures both texture (color intensity) and physical (polarization) inf

Creating Impactful Autonomous Driving Datasets: A Strategic Guide from Research Gap to Benchmark

Model ReleasesDGX agent

arXiv:2607.00710v1 Announce Type: cross Abstract: Well-designed autonomous driving datasets have fundamentally shaped research progress, yet existing literature primarily describes what datasets conta

DeepSeek effect https://x.com/maximelabonne/status/2070867418377818542

Model ReleasesDGX agent

DeepSeek effect https://x.com/maximelabonne/status/2070867418377818542 Fun surprise: DeepSeek used my open-perfectblend dataset to train their new DSpark drafter Time to promote it again! It's an open

DiscoLoop: Looping Discrete Embeddings and Continuous Hidden States for Multi-hop Reasoning

Model ReleasesDGX agent

arXiv:2607.00341v1 Announce Type: cross Abstract: Large language models achieve strong performance on many reasoning tasks when allowed to externalize intermediate steps as Chain-of-Thought (CoT). How

Distributionally Robust Linear Regression With Block Lewis Weights

Model ReleasesDGX agent

arXiv:2607.00252v1 Announce Type: new Abstract: We present an algorithm for the group distributionally robust (GDR) least squares problem. Given m groups, a parameter vector in R^d, and stacked design

Does Your ViT Still Need U-Net for Segmentation?

Model ReleasesDGX agent

arXiv:2607.00223v1 Announce Type: new Abstract: Medical image segmentation is dominated by U-Net-style encoder-decoder architectures. Vision Transformers (ViTs) overcome the limited receptive field of

DriveVA: Video Action Models are Zero-Shot Drivers

Model ReleasesDGX agent

arXiv:2604.04198v2 Announce Type: replace Abstract: Generalization is a central challenge in autonomous driving, as real-world deployment requires robust performance under unseen scenarios, sensor dom

DriveVer: Lightweight Trajectory Evaluator as Test-Time Verifier for Autonomous Driving

Model ReleasesDGX agent

arXiv:2607.00399v1 Announce Type: new Abstract: End-to-end autonomous driving models often encounter performance bottlenecks, as training-time scaling leads to high computational costs and diminishing

DroneFINE: Domain-Aware Parameter-Efficient Fine-Tuning of Vision-Language Detectors for Drone Images

Model ReleasesDGX agent

arXiv:2607.00338v1 Announce Type: new Abstract: Object detection for Unmanned Aerial Vehicles (UAVs) working in open and dynamic environments is a highly challenging task. While Vision-Language Models

Dual-Confidence Contrastive Decoding for Retrieval-Augmented Generation

Model ReleasesDGX agent

arXiv:2607.00570v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) increasingly requires models to answer questions from multiple retrieved documents, where only some sources are rel

Dynamic Bidirectional Pattern Memory: A Production-Scale Empirical Characterisation of Inference-Time Gating in Clinical NLP

Model ReleasesDGX agent

arXiv:2607.00870v1 Announce Type: new Abstract: We study inference-time pattern-memory gating in a production-scale clinical natural language processing (NLP) pipeline. The pipeline pairs a generator

EchoRisk: A Multicentre Echocardiography Dataset and Benchmark for Cardio-Oncology

Model ReleasesDGX agent

arXiv:2607.01039v1 Announce Type: cross Abstract: Therapy-induced cardiotoxicity is the leading non-oncological cause of treatment interruption in breast cancer patients, yet early, automated risk str

EgoGapBench: Benchmarking Egocentric Action Selection in Multi-Agent Scenes

Model ReleasesDGX agent

arXiv:2607.00547v1 Announce Type: cross Abstract: Existing egocentric benchmarks have primarily constructed the egocentric setting from first-person-view data, which makes it difficult to evaluate ego

EgoSafetyBench: A Diagnostic Egocentric Video Benchmark for Evaluating Embodied VLMs as Runtime Safety Guards

Model ReleasesDGX agent

arXiv:2607.00218v1 Announce Type: cross Abstract: Vision-language models (VLMs) are now proposed as runtime safety guards for embodied agents in homes and factories. A deployable guard must catch genu

EmbodimentSemantic: A Spatial Scene-Graph Dataset and Benchmark for Vision-Language Models on Embodied Manipulation Trajectories

Model ReleasesDGX agent

arXiv:2607.00020v1 Announce Type: new Abstract: Spatial grounding remains a key limitation of vision-language-action (VLA) systems for robotic manipulation. While current models can recognize objects

End-to-End Training for Autoregressive Video Diffusion via Self-Resampling

Model ReleasesDGX agent

arXiv:2512.15702v2 Announce Type: replace Abstract: Autoregressive video diffusion models hold promise for world simulation but are vulnerable to exposure bias arising from the train-test mismatch. Wh

Enhanced Vision-Language Models for Diverse Sensor Understanding: Cost-Efficient Optimization and Benchmarking

Model ReleasesDGX agent

arXiv:2412.20750v3 Announce Type: replace Abstract: Large-scale Vision-Language Models (VLMs) have achieved notable progress in aligning visual inputs with text. However, their ability to deeply under

Entropy-Regularized Probabilistic Gates for Sparse Model Discovery in Scarce-Data Federated Learning

Model ReleasesDGX agent

arXiv:2607.00275v1 Announce Type: cross Abstract: Federated Learning (FL) is a distributed machine learning (ML) paradigm with collaboration among multiple clients without sharing data. FL is challeng

EPC: A Standardized Protocol for Measuring Evaluator Preference Dynamics in LLM Agent Systems

Model ReleasesDGX agent

arXiv:2607.00297v1 Announce Type: cross Abstract: When LLM agents use evaluator feedback to adapt their behavior in closed loops, evaluator biases propagate through the agent's strategy distribution -

EVOTS: Evolutionary Transformer Search for Time Series Forecasting

Model ReleasesDGX agent

arXiv:2607.00154v1 Announce Type: cross Abstract: Evolutionary neural architecture design for multivariate time-series forecasting remains underexplored, with most approaches relying on fixed Transfor

ext{Log}_ext{b}Quant: Quantizing Language Models in Logarithmic Space

Model ReleasesDGX agent

arXiv:2607.01127v1 Announce Type: new Abstract: Quantization has become an invaluable tool to reduce memory requirements and inference speed of modern language models, in particular to make them avail

Fable in Claude Code is capable of really amazing things, including for non-coders, but the interface is not really designed for managing 5+…

Model ReleasesDGX agent

Fable in Claude Code is capable of really amazing things, including for non-coders, but the interface is not really designed for managing 5+ hour long autonomous tasks. Really hard to observe what is

FedIA: Importance-Aware Aggregation for Domain-Robust Federated Graph Learning

Model ReleasesDGX agent

arXiv:2509.18171v4 Announce Type: replace Abstract: Federated graph learning (FGL) is a natural paradigm for social-media user graphs, where language communities, regional markets, and service boundar

FlowPath: Learning Data-Driven Manifolds with Invertible Flows for Robust Irregularly-sampled Time Series Classification

Model ReleasesDGX agent

arXiv:2511.10841v3 Announce Type: replace-cross Abstract: Modeling continuous-time dynamics from sparse and irregularly-sampled time series remains a fundamental challenge. Neural controlled different

FLYNN: Robust Neural Network for Robot Navigation using Fly Brain Topology

Model ReleasesDGX agent

arXiv:2607.00025v1 Announce Type: cross Abstract: While deep learning models achieve state-of-the-art performance in complex tasks, they remain brittle when faced with new environments or sensory depr

Foundation Models vs. Radiomics for Lung Computed Tomography: A Benchmark of Feature Extractors, Classification Heads, and Segmentation Choices

Model ReleasesDGX agent

arXiv:2607.01001v1 Announce Type: new Abstract: Radiomics is the established approach for CT-based lung cancer phenotyping, yet comparisons with foundation models rarely isolate contributions of featu

FRAME: Learning the Adaptation Domain with a Mixture of Fractional-Fourier Experts

Model ReleasesDGX agent

arXiv:2607.00162v1 Announce Type: new Abstract: Parameter-efficient fine-tuning (PEFT) reparameterizes weight updates in a fixed basis: low-rank adapters operate in the spatial domain, while a recent

From Signals to Structure: How Memory Architecture Drives Language Emergence in LLM Agents

Model ReleasesDGX agent

arXiv:2607.00233v1 Announce Type: new Abstract: How do two agents invent a shared language from scratch? In a Lewis signaling game, a sender and receiver must coordinate on a code using only their int

From 'Strings' to 'Things' for Personal Knowledge Graphs: Evaluating LLM Triple Extraction for Recommendation Systems

Model ReleasesDGX agent

arXiv:2607.00003v1 Announce Type: cross Abstract: Personal Knowledge Graphs (PKGs) offer a privacy-preserving framework for modeling user preferences, yet constructing them from unstructured, decentra

From Structural Equation Modelling to Double Machine Learning: Robustness Analysis for Survey-Based Research

Model ReleasesDGX agent

arXiv:2607.00512v1 Announce Type: new Abstract: Structural equation modelling (SEM) is widely used in survey-based business and information systems research to assess latent constructs and theory-driv

From Technical Metrics to User Perception: A User Study of a Multimodal Human-Robot Interaction System for Object Detection and Grasping

Model ReleasesDGX agent

arXiv:2607.00530v1 Announce Type: cross Abstract: Improvements in the technical performance of human--robot interaction (HRI) systems do not automatically translate into differences that human users c

FUMO: Prior-Modulated Diffusion for Single Image Reflection Removal

Model ReleasesDGX agent

arXiv:2603.19036v2 Announce Type: replace Abstract: Single image reflection removal (SIRR) is challenging in real scenes, where reflection strength varies spatially and reflection patterns are tightly

FusionFactory: Fusing LLM Capabilities with Multi-LLM Log Data

Model ReleasesDGX agent

arXiv:2507.10540v3 Announce Type: replace Abstract: The rapid advancement of large language models (LLMs) has created a diverse landscape of models, each excelling at different tasks. This diversity d

GAIA: Geometry-Adaptive Operator Learning for Forward and Inverse Problems

Model ReleasesDGX agent

arXiv:2607.01128v1 Announce Type: new Abstract: Operator learning for partial differential equations (PDEs) on arbitrary geometries builds fast neural surrogates for large-scale simulation. Although r

GameDevBench: Evaluating Agentic Capabilities Through Game Development

Model ReleasesDGX agent

arXiv:2602.11103v2 Announce Type: replace Abstract: Despite rapid progress on coding agents, progress on their multimodal counterparts has lagged behind. A key challenge is the scarcity of evaluation

GEAR-Seg: A Grounded Explainable Agent for Reasoning Segmentation and Data Engine

Model ReleasesDGX agent

arXiv:2607.00544v1 Announce Type: new Abstract: Reasoning segmentation requires localizing targets based on complex, implicit queries. Current end-to-end models typically entangle perception and deduc

Geometry-Aware Cross-Height Channel Knowledge Map Prediction for UAV-Assisted Communications With Uncertainty-Guided 3D Sensing

Model ReleasesDGX agent

arXiv:2607.00887v1 Announce Type: new Abstract: Low-altitude Unmanned Aerial Vehicles (UAVs) often need to infer channel knowledge across a range of heights from only sparse observations collected at

GeoSearcher: Anchor-Guided Progressive Reasoning for Remote Sensing Visual Grounding with Process Supervision

Model ReleasesDGX agent

arXiv:2607.01050v1 Announce Type: new Abstract: Recent multimodal large language models (MLLMs) have shown strong cross-modal understanding and coordinate generation abilities in visual grounding. How

Ghost in the Kernel: In-Context Learning with Efficient Transformers via Domain Generalization

Model ReleasesDGX agent

arXiv:2607.00479v1 Announce Type: new Abstract: Transformer-based large models have demonstrated remarkable generalization abilities across different tasks by leveraging a context-aware attention modu

GKDT: General Keypoint Detection Transformer

Model ReleasesDGX agent

arXiv:2607.00752v1 Announce Type: new Abstract: With the emergence of various pre-trained vision and language models, computer vision is shifting from narrow-domain to open-domain recognition. The con

GLM 5.2 DSpark preview is here! ✨ https://huggingface.co/RedHatAI/GLM-5.2-speculator.dspark-preview This is the first DSpark speculator for …

Model ReleasesDGX agent

GLM 5.2 DSpark preview is here! ✨ https://huggingface.co/RedHatAI/GLM-5.2-speculator.dspark-preview This is the first DSpark speculator for a non-DeepSeek frontier model, trained with Speculators and

GMO-E^2DIT: Grounded Multi-Operation Editing for E-Commerce Images

Model ReleasesDGX agent

arXiv:2607.00920v1 Announce Type: new Abstract: Real-world e-commerce image editing often requires multiple, localized, and auditable operations rather than global restyling. This compositional nature

GPTKB v1.5: A Massive Knowledge Base for Exploring Factual LLM Knowledge

Model ReleasesDGX agent

arXiv:2507.05740v2 Announce Type: replace Abstract: Language models are powerful artifacts, yet their factual knowledge is still poorly understood, and inaccessible to ad-hoc browsing and scalable sta

GRACE-RAG: Governed Retrieval Architecture for Canonical Evidence Synthesis, Enabling Lightweight Deployment in Closed-Domain Institutional Settings

Model ReleasesDGX agent

arXiv:2607.00013v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) systems are widely used in institutional question answering settings where responses must be grounded in authorit

Group-Equivariant Poincare Convolutional Networks

Model ReleasesDGX agent

arXiv:2607.00556v1 Announce Type: cross Abstract: While recent advancements like the Poincare ResNet have demonstrated the potential of learning visual representations directly in hyperbolic space, th

GryphOne: Symbol-Aware Masked Diffusion for Structural Refinement in Offline Handwritten Mathematical Expression Recognition

Model ReleasesDGX agent

arXiv:2602.03370v2 Announce Type: replace Abstract: Handwritten mathematical expression recognition (HMER) requires reasoning over diverse symbols and structures, yet autoregressive models struggle wi

GSRQ: Gain-Shape Residual Quantization for Sub-1-bit KV Cache

Model ReleasesDGX agent

arXiv:2607.01065v1 Announce Type: new Abstract: The deployment of Large Language Models (LLMs) with extended context windows is increasingly constrained by the linear growth of Key-Value (KV) cache me

Histopathology Multi-modal Embedding for Pathology Composed Retrieval

Model ReleasesDGX agent

arXiv:2502.07221v4 Announce Type: replace Abstract: To overcome the black-box nature of predictive AI and the hallucination risks of generative models, retrieval-based models offer an interpretable, e

Homogenization of ell_2-Adversarial Training in High-Dimensions: Exact Dynamics under Stochastic Gradient Descent

Model ReleasesDGX agent

arXiv:2607.00207v1 Announce Type: cross Abstract: We develop a framework for analyzing the learning dynamics of ell_2-adversarial training of single-index models on Gaussian mixtures in the high-dimen

huge week of really exciting launches at langchain!

Model ReleasesDGX agent

huge week of really exciting launches at langchain! big week at langchain, with a lot of launches: 1/ OpenWiki - auto generate a wiki of a github repo 2/ two different voice agent tutorials 3/ Harbor

i finally tried hermes agent and the hype is real btw. @NousResearch cooked. been onboarding my young relatives who can't afford Claude, sho…

Model ReleasesDGX agent

i finally tried hermes agent and the hype is real btw. @NousResearch cooked. been onboarding my young relatives who can't afford Claude, showing them how to use $1-5 of tokens to bootstrap hermes and

I gave OpenWiki a spin last night and was very impressed. I've had Claude hand roll wikis for my personal projects/agents, but I threw OpenW…

Model ReleasesDGX agent

I gave OpenWiki a spin last night and was very impressed. I've had Claude hand roll wikis for my personal projects/agents, but I threw OpenWiki at our @subtextdev repo and it handled the whole thing p

I just left the final day of the @aiDotEngineer World's Fair Conference in San Francisco. Kudos to @swyx for putting together a world-class …

Model ReleasesDGX agent

I just left the final day of the @aiDotEngineer World's Fair Conference in San Francisco. Kudos to @swyx for putting together a world-class lineup of speakers and workshops! It really was an invigorat

Identifying and Resolving Pitfalls of Knowledge-Based VQA Benchmarks: Auditing, Repairing, and Augmenting

Model ReleasesDGX agent

arXiv:2607.00159v1 Announce Type: new Abstract: Knowledge-Based Visual Question Answering (KB-VQA) aims to evaluate whether Visual Language Models (VLMs) can retrieve, ground, and reason over external

← Previous
1…109110111112113…377
Next →