AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries89,118
  • Agents7,620
  • Applications5,446
  • Concepts5
  • Hardware1,870
  • Industry6,186
  • Local Ai4,981
  • Model Releases24,181
  • Research20,260
  • Safety13,459
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries89,118
  • Agents7,620
  • Applications5,446
  • Concepts5
  • Hardware1,870
  • Industry6,186
  • Local Ai4,981
  • Model Releases24,181
  • Research20,260
  • Safety13,459
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

Content type
AllBlog
89,118Total entries
1Added by human
89,117Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
64,222 results
Model Releases

v0.32.0

DGX agent

What's Changed New interactive agent experience: running ollama now launches an agent to help you code and delegate work ❯ ollama Ollama 0.32.0 ▸ Chat, Code, & Work (glm-5.2:cloud) Chat with models, c

model-releasesollama-releases
14 Jul 2026
Model Releases

We trained and released DSpark speculators for Kimi-K2.6 and Kimi-K2.7-Code on @huggingface, with native serving support in @vllm_project. A…

DGX agent
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

We trained and released DSpark speculators for Kimi-K2.6 and Kimi-K2.7-Code on @huggingface, with native serving support in @vllm_project. Across six benchmarks in our batch-size-1 evaluation: Kimi-K2

model-releasesclem-delangue--x
13 Jul 2026
Safety

A First-Principles Theory of Slow Thinking and Active Perception

DGX agent

arXiv:2607.08196v1 Announce Type: new Abstract: As part of a series on first-principles modeling of cognitive functions, this paper attempts to provide a mathematical formulation of thinking and perce

safetyarxiv-cs-ai
10 Jul 2026
Model Releases

A Practical Investigation of Training-free Relaxed Speculative Decoding

DGX agent

arXiv:2607.08690v1 Announce Type: cross Abstract: Speculative decoding accelerates sampling from an autoregressive LLM by using a faster auxiliary model to draft tokens which are then verified in para

model-releasesarxiv-cs-ai
10 Jul 2026
Local Ai

A Tool Bottleneck Framework for Clinically-Informed and Interpretable Medical Image Understanding

DGX agent

arXiv:2512.21414v2 Announce Type: replace Abstract: Recent tool-use frameworks powered by vision-language models (VLMs) improve image understanding by grounding model predictions with specialized tool

local-aiarxiv-cs-cv
10 Jul 2026
Local Ai

Asynchronous Federated Continual Segmentation with Evolving Clients and Label Spaces

DGX agent

arXiv:2503.15414v3 Announce Type: replace-cross Abstract: Federated learning seeks to foster collaboration among distributed clients while preserving the privacy of their local data. Traditional feder

local-aiarxiv-cs-cv
10 Jul 2026
Model Releases

Compete Then Collaborate: Frontier AI Teachers Build a Verifiable Curriculum to Improve a Coding Student Beyond Imitation

DGX agent

arXiv:2607.08255v1 Announce Type: new Abstract: Large language models increasingly serve as teachers generating training data for smaller students. Prior multi-teacher knowledge distillation methods m

model-releasesarxiv-cs-ai
10 Jul 2026
Safety

Curriculum Learning for Efficient Chain-of-Thought Distillation via Structure-Aware Masking and GRPO

DGX agent

arXiv:2602.17686v4 Announce Type: replace-cross Abstract: Distilling Chain-of-Thought (CoT) reasoning from large language models into compact student models presents a fundamental challenge: teacher r

safetyarxiv-cs-ai
10 Jul 2026
Model Releases

False Confidence: Automated Labels Confound Fairness Audits in Cervical Spine Segmentation

DGX agent

arXiv:2607.07852v1 Announce Type: cross Abstract: Automated segmentation of cervical-spine MRI is increasingly used in clinical workflows, yet no fairness audit exists for this anatomy. We show that a

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

Hallucination Self-Play: Bootstrapping Reinforced Detector via Evolved Generator

DGX agent

arXiv:2607.07993v1 Announce Type: new Abstract: Identifying faithfulness hallucinations in LLM-generated outputs remains challenging due to the scarcity of high-quality annotated data. Recent work rel

model-releasesarxiv-cs-cl
10 Jul 2026
Model Releases

HumanForge: A Human-Centric Deepfake Video Benchmark with Multi-Agent Forgery Rationales

DGX agent

arXiv:2607.08705v1 Announce Type: new Abstract: Rapid advancements in video diffusion models and temporal editing tools have enabled the generation of highly realistic human-centric videos, posing unp

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

IMProofBench: Benchmarking AI on Research-Level Mathematical Proof Generation

DGX agent

arXiv:2509.26076v2 Announce Type: replace Abstract: As the mathematical capabilities of large language models (LLMs) improve, it becomes increasingly important to evaluate their performance on researc

model-releasesarxiv-cs-cl
10 Jul 2026
Model Releases

LightCrafter: PBR-Conditioned Video Diffusion Refinement for Controllable and Consistent Relighting

DGX agent

arXiv:2607.08016v1 Announce Type: new Abstract: Video relighting requires balancing long-form temporal consistency with a physically grounded understanding of light transport, which depends on accurat

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

LiST: Lipschitz Scaling Training for Robust and Calibrated Neural Networks

DGX agent

arXiv:2607.07745v1 Announce Type: new Abstract: While accuracy, robustness, and calibration are all essential for reliable neural networks, they are often studied separately; developing models that sa

model-releasesarxiv-cs-lg
10 Jul 2026
Model Releases

LSRM: High-Fidelity Object-Centric Reconstruction via Scaled Context Windows

DGX agent

arXiv:2604.05182v2 Announce Type: replace-cross Abstract: We introduce the Large Sparse Reconstruction Model to study how scaling transformer context windows affects feed-forward 3D reconstruction. Al

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Multi-Resolution Feature Stem for Diabetic Retinopathy lesion segmentation

DGX agent

arXiv:2607.08679v1 Announce Type: new Abstract: Diabetic Retinopathy (DR) is a leading cause of preventable blindness worldwide, requiring automated lesion segmentation using deep learning models for

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

Okay, this is winning big time for me right now. Surprised how good GPT-5.6 is at verifiying/advising and all high-level orchestrator capabi…

DGX agent

This post discusses positive experiences with GPT-5.6's capabilities in verification, advisory functions, and high-level orchestration tasks, suggesting the model performs better than expected in thes

model-releasesdair-ai--x
10 Jul 2026
Model Releases

OmniFood-Bench: Evaluating VLMs for Nutrient Reasoning and Personalized Health Advice

DGX agent

arXiv:2607.08423v1 Announce Type: new Abstract: The rapid integration of Large Vision-Language Models (VLMs) into critical infrastructure promises to revolutionize personalized healthcare and dietary

model-releasesarxiv-cs-ai
10 Jul 2026
Research

OpenCoF: Learning to Reason Through Video Generation

DGX agent

arXiv:2607.08763v1 Announce Type: cross Abstract: Reasoning has become a core capability for large models, especially when reliable decisions require understanding logical consequences. Recent video g

researcharxiv-cs-ai
10 Jul 2026
Model Releases

ParamMute: Suppressing Knowledge-Critical FFNs for Faithful Retrieval-Augmented Generation

DGX agent

arXiv:2502.15543v4 Announce Type: replace-cross Abstract: Large language models (LLMs) integrated with retrieval-augmented generation (RAG) have improved factuality by grounding outputs in external ev

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Persuasion Attacks Can Decrease Effectiveness of CoT Monitoring

DGX agent

arXiv:2607.08066v1 Announce Type: new Abstract: Chain-of-thought (CoT) monitoring is a promising safety mechanism for AI agents, based on the premise that visible reasoning traces can surface misalign

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

ReCoLoRA: Spectrum-Aware Recursive Consolidation for Continual LLM Fine-Tuning

DGX agent

arXiv:2607.07719v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning adapts a large language model to one task cheaply, but across a task sequence LoRA-style methods keep stacking low-ran

model-releasesarxiv-cs-ai
10 Jul 2026
Agents

Swapping Faces, Saving Features: A Dual-Purpose Pipeline for Pedestrian Privacy in ITS

DGX agent

arXiv:2607.08402v1 Announce Type: cross Abstract: Large-scale and diverse datasets are needed to train AI models to take real-time decisions for autonomous vehicles (AVs), an intelligent transportatio

agentsarxiv-cs-ai
10 Jul 2026
Safety

Two Axes of LLM Abstention: Answer Correctness and Question Answerability

DGX agent

arXiv:2607.08456v1 Announce Type: cross Abstract: A model should refuse two different things: answers it would get wrong, and questions it should not answer at all, such as unanswerable ones or ones r

safetyarxiv-cs-ai
10 Jul 2026
Model Releases

Uncertainty-gated selection for block-sparse attention

DGX agent

arXiv:2607.07724v1 Announce Type: cross Abstract: Block-sparse attention scales long-context language models by replacing the O(N^2) softmax with a per-query top-k selection over key blocks. This cuto

model-releasesarxiv-cs-cl
10 Jul 2026
Tutorials

VectorizationLLM: Smart Vectorization Based AI Assistant

DGX agent

arXiv:2607.07846v1 Announce Type: new Abstract: VectorizationLLM is a specialized Large Language Model based on Google open-weight LLMs. The model is designed to assist students to learn smart vectori

tutorialsarxiv-cs-ai
10 Jul 2026
Model Releases

CompDiff: Hierarchical Compositional Diffusion for Fair and Zero-Shot Intersectional Medical Image Generation

DGX agent

arXiv:2603.16551v2 Announce Type: replace-cross Abstract: Generative models are increasingly used to augment medical imaging datasets for fairer AI, yet a key assumption often goes unexamined: that ge

model-releasesarxiv-cs-ai
9 Jul 2026
Model Releases

Congrats to our llama cousins 🫡🦙

DGX agent

Congrats to our llama cousins 🫡🦙 Big day for Ollama! When we started, open models and the open source AI ecosystem were in their early days with few believers. Our belief in open source has never wave

model-releasesjerry-liu--x
9 Jul 2026
Model Releases

HAJJv2-CrowdCount: Zero-Shot Benchmark for Dense Crowd Counting

DGX agent

arXiv:2607.07322v1 Announce Type: cross Abstract: Automated crowd counting in Hajj video is difficult not because current models lack capacity, but because the footage violates the assumptions those m

model-releasesarxiv-cs-ai
9 Jul 2026
Model Releases

Optimization-Embedded Active Multi-Fidelity Surrogate Learning for Multi-Condition Airfoil Shape Optimization

DGX agent

arXiv:2603.17057v2 Announce Type: replace-cross Abstract: Active multi-fidelity surrogate modeling is developed for multi-condition airfoil shape optimization to reduce high-fidelity CFD cost while re

model-releasesarxiv-cs-lg
9 Jul 2026
Agents

Physics-Audited Agentic Discovery in Scientific Machine Learning

DGX agent

arXiv:2607.07379v1 Announce Type: new Abstract: In agentic scientific machine learning (SciML), large language model (LLM) agents can discover surrogate models and select one by an automated score, ty

agentsarxiv-cs-ai
9 Jul 2026
Agents

Power and Limitations of Aggregation in Compound AI Systems

DGX agent

arXiv:2602.21556v2 Announce Type: replace Abstract: When designing compound AI systems, a common approach is to query multiple copies of the same model and aggregate the responses to produce a synthes

agentsarxiv-cs-ai
9 Jul 2026
Model Releases

Reinforcement Federated Learning Method Based on Adaptive OPTICS Clustering

DGX agent

arXiv:2306.12859v3 Announce Type: replace Abstract: Federated learning is a distributed machine learning technology, which realizes the balance between data privacy protection and data sharing computi

model-releasesarxiv-cs-lg
9 Jul 2026
Model Releases

Specification Grounding Drives Test Effectiveness for LLM Code

DGX agent

arXiv:2607.06636v1 Announce Type: cross Abstract: Large language models frequently generate code that appears correct on typical inputs yet fails on edge cases, invalid inputs, and other specification

model-releasesarxiv-cs-ai
9 Jul 2026
Safety

The Power of Backdoor Absorption in Community Training

DGX agent

arXiv:2607.06643v1 Announce Type: cross Abstract: Backdoor attacks severely threaten large-scale AI models. When model owners delegate training to external compute providers within a decentralized tra

safetyarxiv-cs-lg
9 Jul 2026
Model Releases

TimEE: End-to-end Time Series Classification via In-Context Learning

DGX agent

arXiv:2607.07500v1 Announce Type: cross Abstract: Time series classification (TSC) is dominated by a two-stage paradigm: train a feature encoder -- either from scratch on the target dataset or via pre

model-releasesarxiv-cs-ai
9 Jul 2026
Model Releases

TRACE-Seg3D: Counterfactual Context Auditing For Robust 3D Glioma Segmentation Under Institutional Shift

DGX agent

arXiv:2607.07038v1 Announce Type: new Abstract: Medical image segmentation models can achieve strong benchmark performance while remaining sensitive to scanner, protocol, and institutional variation.

model-releasesarxiv-cs-cv
9 Jul 2026
Model Releases

Tree-of-Thoughts Reasoning for Text-to-Image In-Context Learning

DGX agent

arXiv:2607.07117v1 Announce Type: cross Abstract: In text-to-image in-context learning (T2I-ICL), a model has to infer a latent compositional pattern from fewshot demonstrations for generating a query

model-releasesarxiv-cs-ai
9 Jul 2026
Model Releases

Try @Grok 4.5!

DGX agent

Try @Grok 4.5! Grok 4.5 is the top non-Anthropic model on AA-Briefcase, combining frontier agentic knowledge work capabilities with leading cost and time-efficiency Yesterday @SpaceXAI released Grok 4

model-releaseselon-musk--x
9 Jul 2026
Model Releases

Unraveling Machine Behavior by Multi-Level Bias Analysis and Detection: Methodology and Application to Computer Vision

DGX agent

arXiv:2607.07236v1 Announce Type: new Abstract: This study investigates the presence and propagation of bias within Neural Networks through a comprehensive multi-level analysis spanning the learned la

model-releasesarxiv-cs-cv
9 Jul 2026
Model Releases

We comprehensively benchmarked GPT-5.6 on document understanding. At a high-level there's no change between GPT-5.6 Sol and GPT-5.5 in terms…

DGX agent

We comprehensively benchmarked GPT-5.6 on document understanding. At a high-level there's no change between GPT-5.6 Sol and GPT-5.5 in terms of performance over tables, text, charts, layout, and more.

model-releasesjerry-liu--x
9 Jul 2026
Model Releases

Auto-DSM Under the Lens: A Black-Box Evaluation Framework for LLM-Based DSM Generation

DGX agent

arXiv:2607.05985v1 Announce Type: new Abstract: This paper presents a black-box evaluation framework to systematically assess the ability of Large Language Models (LLMs) to generate Design Structure M

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

Benchmarking the Robustness of Autonomous Driving to Environmental Illusions: A Lane Perception Perspective

DGX agent

arXiv:2607.05783v1 Announce Type: new Abstract: Environmental illusions (eg., shadows, reflections, and tire marks) are naturally existing yet overlooked phenomena in real-world driving environments.

model-releasesarxiv-cs-cv
8 Jul 2026
Applications

CCBENCH: Assessing LLM Cultural Competence via Implicitly Signaled Norms using Health Queries

DGX agent

arXiv:2607.05405v1 Announce Type: cross Abstract: To interact with users fairly and without stereotyping, AI models must display cultural competency, i.e., the ability to infer and adapt to a user's i

applicationsarxiv-cs-ai
8 Jul 2026
Model Releases

Correct

DGX agent

Correct Grok 4.5 goes public tomorrow. Here’s everything that’s known. The model runs on V9, xAI’s new 1.5 trillion parameter foundation, roughly three times the size of the v8-small architecture behi

model-releaseselon-musk--x
8 Jul 2026
Model Releases

Doomed from the Start: Early Abort of LLM Agent Episodes via a Recall-Controlled Probe Cascade

DGX agent

arXiv:2607.06503v1 Announce Type: new Abstract: Large language model (LLM) agents solving multi-step tasks frequently commit to trajectories that are doomed to fail, yet continue to consume substantia

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

Gemini Enterprise for Education named a Commander in Tambellini StarChart™: 2026 AI Agents for Administrative Efficiency—Agent Platforms

DGX agent

The agentic AI era is here, transforming how higher education institutions innovate, operate, and fundamentally empower learners, faculty, and researchers. AI agents can deliver unprecedented efficien

model-releasesgoogle-cloud-ai
8 Jul 2026
Model Releases

I often use Pinokio to download and work with local open source AI tools/models, and one neat thing is that it exposes them seamlessly to Co…

DGX agent

I often use Pinokio to download and work with local open source AI tools/models, and one neat thing is that it exposes them seamlessly to Code and Codex. GPT-5.6 actually found my local video models a

model-releasesethan-mollick--x
8 Jul 2026
← Previous
1…433434435436437…1338
Next →