AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,555 results
10 Jul 2026

Do You Need a Frontier Model as a Citation Verifier? Benchmarking Rubric LLMs for Deep-Research Source Attribution

Model ReleasesDGX agent

arXiv:2607.08700v1 Announce Type: new Abstract: Reinforcement learning increasingly relies on an LLM judge to score each rubric criterion, and that judge acts as the reward model during training. Befo

DominoTree: Conditional Tree-Structured Drafting with Domino for Speculative Decoding

Model ReleasesDGX agent

arXiv:2607.08642v1 Announce Type: new Abstract: Speculative decoding accelerates LLM inference by drafting several tokens and verifying them in parallel. Block-diffusion drafters such as DFlash produc

Dual-Correlation Hypergraph Network for Unaligned RGBT Video Object Detection and A Large-scale Benchmark


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

arXiv:2607.08191v1 Announce Type: new Abstract: RGB-Thermal (RGBT) Video Object Detection (VOD) has gained significant traction due to its ability to overcome the limitations of conventional RGB-based

Echoes Across Vietnam's Highlands, Delta, and Coast: A Multilingual Corpus for Cham, Khmer, and Tay-Nung

Model ReleasesDGX agent

arXiv:2607.08362v1 Announce Type: new Abstract: Vietnam's ethnic minority languages are almost absent from the field of Natural Language Processing (NLP), and the challenge goes beyond data scarcity:

Equivariant Quantum Clustering with Differential Privacy: Parameter-Efficient Privacy-Preserving Analysis Across Heterogeneous Sensitive Datasets

Model ReleasesDGX agent

arXiv:2607.08092v1 Announce Type: cross Abstract: Privacy-preserving clustering is critical for analyzing sensitive data in healthcare, cybersecurity, and enterprise applications, where maintaining da

Evaluating the Generalizability of Foundation Models for Extreme Environmental Events: Case Study of California Wildfire PM2.5

Model ReleasesDGX agent

arXiv:2607.07951v1 Announce Type: new Abstract: Wildfire smoke events produce extreme PM_{2.5} concentrations that pose severe public health risks, yet forecasting rare, hazardous-level spikes remains

FabriVLA: A Lightweight Vision-Language-Action Model for Precise Multi-Task Manipulation

Model ReleasesDGX agent

arXiv:2607.08575v1 Announce Type: new Abstract: We present FabriVLA, a lightweight Vision-Language-Action model for Precise Multi-Task Manipulation. FabriVLA combines an InternVL3.5 vision-language ba

False Confidence: Automated Labels Confound Fairness Audits in Cervical Spine Segmentation

Model ReleasesDGX agent

arXiv:2607.07852v1 Announce Type: cross Abstract: Automated segmentation of cervical-spine MRI is increasingly used in clinical workflows, yet no fairness audit exists for this anatomy. We show that a

Fine-tune NVIDIA Nemotron 3 models with Amazon SageMaker AI serverless model customization

Model ReleasesDGX agent

In this post, we explore what makes the Nemotron 3 architecture unique, walk through the fine-tuning techniques available, and show you step-by-step how to get started with serverless customization us

folks who’ve tried Claude Design — how was the experience? what works well and what could be better?

Model ReleasesDGX agent

Claude Design appears to be a design-focused tool or feature within Claude that users can interact with; this post solicits feedback from people who have tried it, asking them to evaluate what aspects

Formal Mechanisms for Market Stability in Self-Interested Agent Societies: A Marketplace Simulation Study

Model ReleasesDGX agent

arXiv:2607.08652v1 Announce Type: new Abstract: Self-interested agents, left unconstrained, tend toward defection in repeated social dilemmas, causing cooperative gains from trade to collapse. This pa

Frontier and Center: Who evaluates the evaluations?

Model ReleasesDGX agent

Editor’s note: Some of the most interesting questions in AI are being asked by information theoreticians, around how to provide context to an emerging class of AI agents. A few weeks ago, we waded int

Functional and Secure Code Generation with Task Vectors

Model ReleasesDGX agent

arXiv:2607.07881v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used for code generation, but they struggle to generate functional code free of security vulnerabilities

Generalization Theory for Through-the-Wall Radar Human Activity Recognition

Model ReleasesDGX agent

arXiv:2607.08144v1 Announce Type: cross Abstract: Through-the-wall radar (TWR) human activity recognition (HAR) is important for non-line-of-sight indoor sensing, security monitoring, and emergency re

Generative Action Tell-Tales: Assessing Human Motion in Synthesized Videos

Model ReleasesDGX agent

arXiv:2512.01803v3 Announce Type: replace Abstract: Despite rapid advances in video generative models, robust metrics for evaluating visual and temporal correctness of complex human actions remain elu

Geometry and Gradient-based Partitioning for Panoramic Outdoor Reconstruction

Model ReleasesDGX agent

arXiv:2607.08769v1 Announce Type: new Abstract: Scaling 3D Gaussian Splatting (3DGS) to large outdoor scenes is costly in both data acquisition and computation. Adopting panoramic images with equirect

Glad to share what I've been building in public. A lightweight AI news feed. We also finetuned a model to write new stories better than Clau…

Model ReleasesDGX agent

Glad to share what I've been building in public. A lightweight AI news feed. We also finetuned a model to write new stories better than Claude or GPT with extensive prompting (believe me not for a lac

GPT-5.6 is a major step forward for health intelligence. Across the lineup, we’re delivering stronger performance at lower cost: GPT-5.6 Lun…

Model ReleasesDGX agent

GPT-5.6 is a major step forward for health intelligence. Across the lineup, we’re delivering stronger performance at lower cost: GPT-5.6 Luna outperforms GPT-5.5 at its highest reasoning setting while

GPT-5.6 Terra and Sol are now available in Perplexity and Perplexity Computer.

Model ReleasesDGX agent

Perplexity has released two new AI models, GPT-5.6 Terra and Sol, which are now available for use in Perplexity's search platform and Perplexity Computer tool. These models represent updates to Perple

Graph-Regularized Deep Learning for EEG-Based Emotion Recognition with Psychologically-Grounded Label Structure

Model ReleasesDGX agent

arXiv:2607.07773v1 Announce Type: cross Abstract: EEG-based emotion recognition is critical for mental health monitoring and affective brain-computer interfaces, yet existing deep learning approaches

great viz by @BraceSproul to put this in perspective. the really important part of this is that OpenWiki is additive!

Model ReleasesDGX agent

great viz by @BraceSproul to put this in perspective. the really important part of this is that OpenWiki is additive! OpenWiki general purpose memory is meant to be complementary to codex/claude code

GROK 4.5 LEADS ON REAL PROFESSIONAL WORK BENCHMARK New data from Snorkel shows Grok 4.5 outperforming other frontier models on real-world pr…

Model ReleasesDGX agent

GROK 4.5 LEADS ON REAL PROFESSIONAL WORK BENCHMARK New data from Snorkel shows Grok 4.5 outperforming other frontier models on real-world professional tasks. On their GDPval+ benchmark (expert-created

Grok Build

Model ReleasesDGX agent

Grok Build Grok 4.5 with Grok Build just ranked #1 on the SWE-Atlas-QnA benchmark with a score of 84 That puts it level with GPT-5.6 (max) Codex and ahead of Claude Code Fable 5 (max), Opus 4.8 (max),

Grok doesn’t give up

Model ReleasesDGX agent

Grok doesn’t give up Grok 4.5 is the most persistent agent model we've tested. Here's one example: In one of our evals, we asked 3 models (GPT-5.5, GLM-5.2 and Grok 4.5) to audit a GitHub repo for har

Grok is closing the loop on real-world use cases

Model ReleasesDGX agent

Grok is closing the loop on real-world use cases @OpenAI released 3 new models yesterday and we immediately tested it on our internal benchmark. All three outperforming gpt-5.5 but @SpaceXAI still the

Hallucination Self-Play: Bootstrapping Reinforced Detector via Evolved Generator

Model ReleasesDGX agent

arXiv:2607.07993v1 Announce Type: new Abstract: Identifying faithfulness hallucinations in LLM-generated outputs remains challenging due to the scarcity of high-quality annotated data. Recent work rel

How Deutsche Telekom is rewiring telecommunications with AI

Model ReleasesDGX agent

Deutsche Telekom is implementing AI technologies to modernize and optimize its telecommunications infrastructure and operations. The article likely discusses specific applications of AI in areas such

How to get 100% consistent product ads from one seedance 2.0 generation all directed from one chat with the hashtag Comfy MCP. @hellorob did…

Model ReleasesDGX agent

How to get 100% consistent product ads from one seedance 2.0 generation all directed from one chat with the hashtag Comfy MCP. @hellorob didn't let the agent improvise a pipeline. He pointed it at his

Hugging Face Gemma Challenge results are in! 📈 Over 6 days, more than 100 AI agents and humans collaborated to make Gemma 4 inference 5x fa…

Model ReleasesDGX agent

Hugging Face Gemma Challenge results are in! 📈 Over 6 days, more than 100 AI agents and humans collaborated to make Gemma 4 inference 5x faster on a single NVIDIA A10G GPU. - Fastest result: 491.8 TPS

HumanForge: A Human-Centric Deepfake Video Benchmark with Multi-Agent Forgery Rationales

Model ReleasesDGX agent

arXiv:2607.08705v1 Announce Type: new Abstract: Rapid advancements in video diffusion models and temporal editing tools have enabled the generation of highly realistic human-centric videos, posing unp

ICDAR 2026 HIPE-OCRepair Competition on LLM-Assisted OCR Post-Correction for Historical Documents

Model ReleasesDGX agent

arXiv:2607.08143v1 Announce Type: cross Abstract: We present the results of HIPE-OCRepair-2026, an ICDAR competition on LLM-assisted OCR post-correction of historical documents. OCR post-correction re

Ideas Have Genomes: Benchmarking Scientific Lineage Reasoning and Lineage-Grounded Idea Generation

Model ReleasesDGX agent

arXiv:2607.08758v1 Announce Type: new Abstract: Scientific ideas rarely start from a blank page. They inherit mechanisms, repair known limitations, and recombine pieces of earlier work, much like biol

IFAR: Multi-Perspective and Multi-Level Causal Discovery with LLMs

Model ReleasesDGX agent

arXiv:2409.05559v2 Announce Type: replace Abstract: Large language models (LLMs) have developed rapidly, and their reasoning capabilities have become a hot research topic. However, there is still limi

IMProofBench: Benchmarking AI on Research-Level Mathematical Proof Generation

Model ReleasesDGX agent

arXiv:2509.26076v2 Announce Type: replace Abstract: As the mathematical capabilities of large language models (LLMs) improve, it becomes increasingly important to evaluate their performance on researc

Improving Ad-hoc Search Effectiveness for Conversational Information Retrieval via Model Merging

Model ReleasesDGX agent

arXiv:2607.08540v1 Announce Type: cross Abstract: Conversational information retrieval is challenging since it requires the consideration of the conversation history which potentially gives rise to to

Infinity-Parser2 Technical Report

Model ReleasesDGX agent

arXiv:2607.07836v1 Announce Type: new Abstract: We present Infinity-Parser2, a large multimodal model that couples a controllable data-synthesis pipeline with multi-task reinforcement learning for end

Jet-Long: Efficient Long-Context Extension with Dynamic Bifocal RoPE

Model ReleasesDGX agent

arXiv:2607.07740v1 Announce Type: cross Abstract: Modern LLMs are increasingly deployed in long-context applications such as retrieval-augmented generation, repository-level coding, and agentic workfl

Joint Bayesian Parameter and Model Order Estimation for Low-Rank Probability Mass Tensors

Model ReleasesDGX agent

arXiv:2410.06329v4 Announce Type: replace-cross Abstract: Obtaining a reliable estimate of the joint probability mass function (PMF) of a set of random variables from observed data is a significant ob

Joint Discrete-Continuous Flow Matching for Open-Vocabulary Inverse Design of Multilayer Optical Coatings

Model ReleasesDGX agent

arXiv:2607.08392v1 Announce Type: cross Abstract: Amortized neural inverse design typically remains closed-world: component choices are fixed vocabulary tokens, coordinate grids are frozen at training

Key to this strategy is deterrence, 'Mutually Assured Compute Destruction' which gets its own section. It doesn't mention the generalization…

Model ReleasesDGX agent

Key to this strategy is deterrence, 'Mutually Assured Compute Destruction' which gets its own section. It doesn't mention the generalization Mutually Assured AI Malfunction (MAIM) from Schmidt, Wang,

KronQ: LLM Quantization via Kronecker-Factored Hessian

Model ReleasesDGX agent

arXiv:2607.07964v1 Announce Type: new Abstract: Post-training quantization (PTQ) is a widely adopted technique for compressing large language models (LLMs) without retraining. Existing second-order PT

🚀langchain launches this week: all about open source models and memory! First: open source models. We partnered with @NVIDIAAI to launch a …

Model ReleasesDGX agent

🚀langchain launches this week: all about open source models and memory! First: open source models. We partnered with @NVIDIAAI to launch a NemoClaw DeepAgents blueprint. This pairs Deep Agents (our op

Less Data, Faster Convergence: Goal-Driven Data Optimization for Multimodal Instruction Tuning

Model ReleasesDGX agent

arXiv:2603.12478v2 Announce Type: replace Abstract: Multimodal instruction tuning is often compute-inefficient because training budgets are spread across large mixed image-video pools whose utility is

LEXIC: Lightweight Eye-tracking eXtension via Injected Complexity

Model ReleasesDGX agent

arXiv:2607.08152v1 Announce Type: cross Abstract: On the recent EyeBench benchmark, predicting reading comprehension from eye movements exposes a stark gap: text-aware models using pretrained language

LightCrafter: PBR-Conditioned Video Diffusion Refinement for Controllable and Consistent Relighting

Model ReleasesDGX agent

arXiv:2607.08016v1 Announce Type: new Abstract: Video relighting requires balancing long-form temporal consistency with a physically grounded understanding of light transport, which depends on accurat

Like the Claude app it imitates, the new ChatGPT 'Super App', which merges Codex with ChatGPT, is a tangle of toggles and strange UI decisions (M.G. Siegler/Spyglass)

Model ReleasesDGX agent

M.G. Siegler / Spyglass: Like the Claude app it imitates, the new ChatGPT “Super App”, which merges Codex with ChatGPT, is a tangle of toggles and strange UI decisions — Even if it's ultimately the ri

Linear Attention Architectures: Mechanisms, Trade-offs, and Cross-Layer Routing

Model ReleasesDGX agent

arXiv:2607.07953v1 Announce Type: cross Abstract: Self-attention lets each token retrieve information from the full context, but its quadratic cost in sequence length limits training and inference at

LiST: Lipschitz Scaling Training for Robust and Calibrated Neural Networks

Model ReleasesDGX agent

arXiv:2607.07745v1 Announce Type: new Abstract: While accuracy, robustness, and calibration are all essential for reliable neural networks, they are often studied separately; developing models that sa

LlamaSeg: Image Segmentation via Autoregressive Mask Generation

Model ReleasesDGX agent

arXiv:2505.19422v2 Announce Type: replace Abstract: We present extbf{LlamaSeg}, a visual autoregressive framework that unifies multiple image segmentation tasks via natural language instructions. By r

LSRM: High-Fidelity Object-Centric Reconstruction via Scaled Context Windows

Model ReleasesDGX agent

arXiv:2604.05182v2 Announce Type: replace-cross Abstract: We introduce the Large Sparse Reconstruction Model to study how scaling transformer context windows affects feed-forward 3D reconstruction. Al

LUMI: Tokenizer-Agnostic LLM-Based Lossless Image Compression

Model ReleasesDGX agent

arXiv:2607.08221v1 Announce Type: new Abstract: Large language model (LLM)-based lossless image compression methods typically represent pixel data through the native text interface of a pretrained mod

MetaHGNIE: Meta-Path Induced Hypergraph Contrastive Learning in Heterogeneous Knowledge Graphs

Model ReleasesDGX agent

arXiv:2512.12477v2 Announce Type: replace Abstract: Estimating node importance in heterogeneous knowledge graphs is a fundamental problem underlying recommendation, search, and knowledge decision syst

Mixture of Enhanced-View Experts for Multi-Query Vehicle ReID and A Large-Scale Benchmark

Model ReleasesDGX agent

arXiv:2607.08085v1 Announce Type: new Abstract: Multi-query vehicle ReID aims to leverage complementary information from diverse views for robust feature learning. However, current methods suffer from

MSRNet: A Multi-Scale Recursive Network for Camouflaged Object Detection

Model ReleasesDGX agent

arXiv:2511.12810v2 Announce Type: replace-cross Abstract: Camouflaged object detection is an emerging and challenging computer vision task that requires identifying and segmenting objects that blend s

Multi-Resolution Feature Stem for Diabetic Retinopathy lesion segmentation

Model ReleasesDGX agent

arXiv:2607.08679v1 Announce Type: new Abstract: Diabetic Retinopathy (DR) is a leading cause of preventable blindness worldwide, requiring automated lesion segmentation using deep learning models for

New research from Meta. (bookmark it) It's on how to fix agents that forget previously made decisions. It's well know that long-horizon agen…

Model ReleasesDGX agent

New research from Meta. (bookmark it) It's on how to fix agents that forget previously made decisions. It's well know that long-horizon agents keep forgetting decisions they already made. Meta researc

@nickfrosst Full podcast > https://www.youtube.com/watch?v=pc4vT8xcSaY

Model ReleasesDGX agent

Nick Frost, likely a notable figure in AI or machine learning, discusses his work and insights in a full-length podcast available on YouTube. The podcast was shared by Cohere, an AI company, suggestin

Not everyone will succeed in the model business. 'You've seen a lot of companies spend a huge amount on compute and then not end up with a g…

Model ReleasesDGX agent

Cohere warns that computational investment alone does not guarantee success in building AI models, emphasizing that many companies spend substantial resources on computing infrastructure without achie

Okay, this is winning big time for me right now. Surprised how good GPT-5.6 is at verifiying/advising and all high-level orchestrator capabi…

Model ReleasesDGX agent

This post discusses positive experiences with GPT-5.6's capabilities in verification, advisory functions, and high-level orchestration tasks, suggesting the model performs better than expected in thes

OmniFood-Bench: Evaluating VLMs for Nutrient Reasoning and Personalized Health Advice

Model ReleasesDGX agent

arXiv:2607.08423v1 Announce Type: new Abstract: The rapid integration of Large Vision-Language Models (VLMs) into critical infrastructure promises to revolutionize personalized healthcare and dietary

← Previous
1…8687888990…376
Next →