AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,841 results
Model Releases

SealQA: Raising the Bar for Reasoning in Search-Augmented Language Models

DGX agent

arXiv:2506.01062v4 Announce Type: replace Abstract: We introduce SealQA, a new challenge benchmark for evaluating SEarch-Augmented Language models on fact-seeking questions where web search yields con

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Self-Supervised Foundation Model for Calcium-imaging Population Dynamics

DGX agent

arXiv:2604.04958v2 Announce Type: replace-cross Abstract: Recent work suggests that large-scale, multi-animal modeling can significantly improve neural recording analysis. However, for functional calc

model-releasesarxiv-cs-ai
10 Apr 2026
Research

Squeeze Evolve: Unified Multi-Model Orchestration for Verifier-Free Evolution

DGX agent

arXiv:2604.07725v1 Announce Type: cross Abstract: We show that verifier-free evolution is bottlenecked by both diversity and efficiency: without external correction, repeated evolution accelerates col

researcharxiv-cs-cl
10 Apr 2026
Model Releases

Tool Retrieval Bridge: Aligning Vague Instructions with Retriever Preferences via Bridge Model

DGX agent

arXiv:2604.07816v1 Announce Type: new Abstract: Tool learning has emerged as a promising paradigm for large language models (LLMs) to address real-world challenges. Due to the extensive and irregularl

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

VSAS-BENCH: Real-Time Evaluation of Visual Streaming Assistant Models

DGX agent

arXiv:2604.07634v1 Announce Type: new Abstract: Streaming vision-language models (VLMs) continuously generate responses given an instruction prompt and an online stream of input frames. This is a core

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

We pit LlamaParse against frontier models (Opus 4.6, Gemini 3.1 Pro, GPT-5.4) in a live OCR arena. ICYMI: the full workshop is on Youtube! F…

DGX agent

We pit LlamaParse against frontier models (Opus 4.6, Gemini 3.1 Pro, GPT-5.4) in a live OCR arena. ICYMI: the full workshop is on Youtube! Frontier VLMs are getting quite good at visual understanding,

model-releasesjerry-liu--x
10 Apr 2026
Tutorials

What do Language Models Learn and When? The Implicit Curriculum Hypothesis

DGX agent

arXiv:2604.08510v1 Announce Type: new Abstract: Large language models (LLMs) can perform remarkably complex tasks, yet the fine-grained details of how these capabilities emerge during pretraining rema

tutorialsarxiv-cs-cl
10 Apr 2026
Model Releases

Google’s Gemini AI can answer your questions with 3D models and simulations

DGX agent

Google's latest upgrade for Gemini will allow the chatbot to generate interactive 3D models and simulations in response to your questions. With the new feature, you may see options to rotate the AI-ge

model-releasesthe-verge-ai
9 Apr 2026
Model Releases

Google's Gemma 4 is pretty wild. You can now run it locally with OpenClaw in 3 steps. 1. Install Ollama 2. Pull Gemma 4 model 3. Launch Open…

DGX agent

Google's Gemma 4 is pretty wild. You can now run it locally with OpenClaw in 3 steps. 1. Install Ollama 2. Pull Gemma 4 model 3. Launch OpenClaw with Gemma as the backend Private local AI agents in mi

model-releasesollama--x
8 Apr 2026
Model Releases

🏆 Qwen3.6-Plus just swept OpenRouter’s Daily, Weekly & Trending charts! The trial has wrapped. The model is fully live & production-ready. …

DGX agent

🏆 Qwen3.6-Plus just swept OpenRouter’s Daily, Weekly & Trending charts! The trial has wrapped. The model is fully live & production-ready. 📉 Lower latency | 🔥 Top-tier reasoning | 💰 Unmatched $/token

model-releasesqwen--x
8 Apr 2026
Model Releases

The new model from Meta is already looking like a disappointment: overoptimized for public benchmark numbers at the detriment of everything …

DGX agent

The new model from Meta is already looking like a disappointment: overoptimized for public benchmark numbers at the detriment of everything else. Knowing how to evaluate models in a way that correlate

model-releasesfrancois-chollet--x
8 Apr 2026
Research

First-order friction models with bristle dynamics: lumped and distributed formulations

DGX agent

arXiv:2602.09429v3 Announce Type: replace-cross Abstract: Dynamic models, particularly rate-dependent models, have proven effective in capturing the key phenomenological features of frictional process

researcharxiv-cs-ro
13 Aug 2026
Model Releases

How China-Origin Vision-Language Models Move from Refusal to Reframing in State Alignment

DGX agent

arXiv:2608.11816v1 Announce Type: cross Abstract: State-aligned distortion has been documented in China-origin text-based large language models (LLMs), but whether, and in what form, it arises in mult

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

Localizing Safety Alignment: MLP Layers and Mid-Network Blocks Encode Refusal Behavior in Large Language Models

DGX agent

arXiv:2608.11583v1 Announce Type: new Abstract: Safety alignment in large language models is often treated as a distributed property of the entire network, yet its practical brittleness suggests that

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

Retrofitting Recurrent Depth into a Pretrained Language Model: Installation, Extrapolation, Transfer, and Retention at Two Parameter Budgets

DGX agent

arXiv:2608.11233v1 Announce Type: cross Abstract: A dense, pretrained language model can be retrofitted with recurrent depth and learn an iterative latent transition that persists after outcome-only a

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

Four small architecture decisions can cost up to 47% of a model's long-context performance. New research from Ai2, Carnegie Mellon, and the …

DGX agent

Four small architecture decisions can cost up to 47% of a model's long-context performance. New research from Ai2, Carnegie Mellon, and the University of Washington isolates them. Normalization, GQA,

model-releasesdair-ai--x
12 Aug 2026
Model Releases

More Accurate, Less Human: Gestalt Grouping in Vision Models

DGX agent

arXiv:2608.10195v1 Announce Type: new Abstract: Human vision organizes what it sees into wholes: same-colored points group into series, similar marks cohere into categories, and shapes complete into r

model-releasesarxiv-cs-cv
12 Aug 2026
Model Releases

Optimal Stopping of Self-Refining Foundation Models

DGX agent

arXiv:2608.10729v1 Announce Type: cross Abstract: Foundation models can improve their outputs through a self-refinement process driven by external feedback. In this process, the model is embedded in a

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

ProTAGAD: A Foundation Model for TAG Anomaly Detection with Decoupled Topological and Textual Prototypes

DGX agent

arXiv:2608.10699v1 Announce Type: cross Abstract: Text-Attributed Graphs (TAGs), endowed with abundant textual content along with topological structures, have emerged as a versatile backbone for real-

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

Qwen3.8-2.4T-A95B is now live on Fireworks with Day-0 support! This 2.4T parameter MoE model is built for autonomous agents, heavy coding, l…

DGX agent

Qwen3.8-2.4T-A95B is now live on Fireworks with Day-0 support! This 2.4T parameter MoE model is built for autonomous agents, heavy coding, large context windows, and is ideal for coding and agentic pe

model-releasesfireworks-ai--x
12 Aug 2026
Model Releases

TemMed-Bench: Evaluating Temporal Medical Image Reasoning in Vision-Language Models

DGX agent

arXiv:2509.25143v2 Announce Type: replace-cross Abstract: Existing medical reasoning benchmarks for vision-language models primarily focus on analyzing a patient's condition based on an image from a s

model-releasesarxiv-cs-cl
12 Aug 2026
Model Releases

The Truth Stays in the Family: Enhancing Contextual Grounding via Inherited Truthful Heads in Model Lineages

DGX agent

arXiv:2606.15821v2 Announce Type: replace-cross Abstract: Recent advances in large language models (LLMs) have produced many specialized multimodal LLMs (MLLMs) that share common foundational LLMs, fo

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

AnyCamVLA: Zero-Shot Camera Adaptation for Viewpoint Robust Vision-Language-Action Models

DGX agent

arXiv:2603.05868v2 Announce Type: replace Abstract: Despite remarkable progress in Vision-Language-Action models (VLAs) for robot manipulation, these large pre-trained models require fine-tuning to be

model-releasesarxiv-cs-ro
11 Aug 2026
Model Releases

Do All LLMs Know When They're Being Harmful? A Reproducibility Study of Latent-Space Safety Probes Across Model Families

DGX agent

arXiv:2608.08029v1 Announce Type: cross Abstract: Khatri et al. (2026) [DOI: 10.1109/DSN-W70714.2026.00027] show that lightweight MLP probes on final-layer activations of a single 8B model (LLaMA-3.1-

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Evo-Bench: Can Language Models Improve Agent Harness?

DGX agent

arXiv:2608.09096v1 Announce Type: new Abstract: Large Language Models (LLMs) have driven rapid progress in autonomous agents, yet standard evaluations remain confined to static task solving. An emergi

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

From Values to Benchmarks: Evaluating Large Language Models for Governmental Use in Dutch

DGX agent

arXiv:2608.09925v1 Announce Type: cross Abstract: Large language models are increasingly being deployed in governmental settings, yet few existing evaluation frameworks jointly reflect the values of p

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Looking for a faster specialized model for your Agent Work? @nvidia Nemotron 3.5 Lightning (30B MoE, 3B active params) is now live on Firewo…

DGX agent

Looking for a faster specialized model for your Agent Work? @nvidia Nemotron 3.5 Lightning (30B MoE, 3B active params) is now live on Fireworks. It’s distilled from NVIDIA Nemotron 3 Ultra to be your

model-releasesfireworks-ai--x
11 Aug 2026
Safety

On the use of foundation models in cognitive science

DGX agent

arXiv:2608.07812v1 Announce Type: new Abstract: A host of recent studies have evaluated the cognitive and developmental alignment of Foundation Models (FMs). These investigations include evaluations o

safetyarxiv-cs-cl
11 Aug 2026
Model Releases

Other active promotions: - Free models: Solar Pro 4 (1 week), Hy3, Step 3.7 Flash, Laguna S and XS - 90% off DeepSeek V4 Flash for ~2 more d…

DGX agent

Nous Research has extended its 20 % discount on all models—including high‑end frontier options—throughout the Nous Portal for an additional two weeks (until the end of April). Free model trials such a

model-releasesnous-research--x
11 Aug 2026
Research

Second Order Drifting Models

DGX agent

arXiv:2608.07924v1 Announce Type: cross Abstract: Drifting models are a recent class of one-step generative models that evolve the model distribution during training using a predefined sample-based dr

researcharxiv-cs-ai
11 Aug 2026
Model Releases

SurgWMBench: A Vision-Based Benchmark for World-Modeling Surgical Instrument Motion Planning

DGX agent

arXiv:2608.08070v1 Announce Type: new Abstract: Reliable surgical planning requires models that move beyond recognizing the current surgical step or imitating expert demonstrations, and instead antici

model-releasesarxiv-cs-ro
11 Aug 2026
Model Releases

b10344

DGX agent

model: add MTP support for Nemotron model (#26725) model: add MTP support for Nemotron Nano model model: add mtp_flags for nemotron model address review comments Website: https://llama.app macOS/iOS:

model-releasesllama-cpp-releases
10 Aug 2026
Tutorials

Decoupling Intention from Trajectory: A Representational Deduction Framework for World Action Models

DGX agent

arXiv:2608.06994v1 Announce Type: cross Abstract: World Action Models (WAMs) aim to construct a unified architecture capable of understanding world state evolution and guiding to generative motion pla

tutorialsarxiv-cs-ai
10 Aug 2026
Safety

Genotypic Triggers: Exposing Pharmacogenomic Blind Spots via Host-Specific Backdoors in Generative Antimicrobial Peptide Models

DGX agent

arXiv:2608.06779v1 Announce Type: cross Abstract: Large Language Models (LLMs) have accelerated drug discovery, particularly in the automated design of antimicrobial peptides (AMPs). However, current

safetyarxiv-cs-ai
10 Aug 2026
Model Releases

MirrorWorld: Taming Video Diffusion Models for Mirror Reflection Generation

DGX agent

arXiv:2608.07463v1 Announce Type: new Abstract: Recent advances in video diffusion models (VDMs) have enabled high-fidelity video synthesis. However, generating mirror reflections remains challenging

model-releasesarxiv-cs-cv
10 Aug 2026
Model Releases

Model Confidence Under Answer-Preserving Attacks: An Informativeness-Manipulability Frontier

DGX agent

arXiv:2608.06571v1 Announce Type: cross Abstract: Deployed vision-language systems often gate their answers on confidence, making confidence robustness relevant to oversight. We study confidence reado

model-releasesarxiv-cs-cl
10 Aug 2026
Safety

SoRoMoX: Fast, Differentiable, and Parallelizable Soft Robot Models

DGX agent

arXiv:2608.06650v1 Announce Type: cross Abstract: Reduced-order models based on Cosserat-rod theory are now well established, and modeling theory is no longer the primary bottleneck in soft-robot cont

safetyarxiv-cs-ai
10 Aug 2026
Model Releases

model: support Longcat-Flash (need testing) by ngxson · Pull Request #19182 · ggml-org/llama.cpp

DGX agent

This PR should be ready for testing now. I tested with a very small (8B params) sub-model extracted from the original one. Appreciate if someone can test with the bigger model. GGUF(for testing) from

model-releasesr-localllama
8 Aug 2026
Research

Diff-VF: Training-free High-quality Long Video Generation via Diffusion Model

DGX agent

arXiv:2608.05976v1 Announce Type: new Abstract: Recently, diffusion models have made great progress in video generation. However, most existing video diffusion models are trained with short videos, an

researcharxiv-cs-cv
7 Aug 2026
Agents

How cheap models changed multi-agent economics

DGX agent

Orchestrator-executor just became the smart default for production agents: an expensive model plans, cheap models execute, and cost per completed task decides the roster. The post How cheap models cha

agentsarize-ai
7 Aug 2026
Model Releases

Zero-Shot Multi-Disease Labeling of Chest, Abdomen, and Pelvis CT Reports Using Open-Weight Large Language Models: The Effect of Labeling Conventions

DGX agent

arXiv:2506.03259v3 Announce Type: replace Abstract: Purpose: To compare five lightweight open-weight large language models (LLMs) with a rule-based algorithm (RBA) and fine-tuned RadBERT for zero-shot

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

BrainBench: Benchmarking Large Language Models for Comprehensive EEG Understanding

DGX agent

arXiv:2608.04156v1 Announce Type: new Abstract: Electroencephalography (EEG) analysis extends beyond assigning predefined labels to recordings; it requires workflows connecting natural-language instru

model-releasesarxiv-cs-ai
6 Aug 2026
Research

DASyR-LLM: Domain-Aware Symbolic Regression with LLMs for Kinetic Model Discovery

DGX agent

arXiv:2608.05120v1 Announce Type: new Abstract: Kinetic model discovery is a central challenge in chemical engineering, as accurate rate expressions are essential for understanding and controlling che

researcharxiv-cs-lg
6 Aug 2026
Model Releases

NuclearDiffusion: Text-to-Image Foundation Models for Learning Nuclear Energy Concepts

DGX agent

arXiv:2608.04030v1 Announce Type: cross Abstract: Generative artificial intelligence (AI) has transformed text-to-image synthesis, yet its ability to represent specialized engineering domains remains

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

SciCode-Verified: How Benchmark Defects Underestimated the Scientific-Coding Ability of Language Models

DGX agent

arXiv:2608.04975v1 Announce Type: cross Abstract: SciCode is the standard measure of the scientific-coding ability of language models: research-level problems that demand both frontier scientific theo

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

WorldCycle: Self-Verifiable Reinforcement Learning for Long-Horizon Video World Models

DGX agent

arXiv:2608.04964v1 Announce Type: new Abstract: Interactive video world models are essential for long-horizon planning and exploration, yet they suffer from compounding errors. Post-training methods s

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

CrossScope: A Role-Asymmetric World Model for Joint Dual-Scope Surgical Video Prediction

DGX agent

arXiv:2608.03211v1 Announce Type: new Abstract: Visual world models typically learn future dynamics from a single observation stream, limiting their ability to model cooperative systems with multiple

model-releasesarxiv-cs-cv
5 Aug 2026
Applications

Modeling Scientific Experiment Scenes: Dataset and Model

DGX agent

arXiv:2608.02892v1 Announce Type: new Abstract: Scene Graph Generation (SGG) is fundamental to structured visual understanding, yet existing benchmarks focus mainly on daily life images and overlook s

applicationsarxiv-cs-cv
5 Aug 2026
← Previous
1…3233343536…1247
Next →