AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlog
88,419Total entries
1Added by human
88,418Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,638 results
Safety

Covariance-Boosted Gaussian Processes for Spatiotemporal Irregularities

DGX agent

arXiv:2607.23018v1 Announce Type: cross Abstract: Nonstationary Gaussian process (GP) models are powerful tools for capturing input-dependent variability by adapting to observed data. However, with li

safetyarxiv-cs-lg
28 Jul 2026
Safety

Data Pyramid for Embodied Manipulation

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2607.24744v1 Announce Type: cross Abstract: Multimodal foundation models learned to see and to speak by consuming the whole internet. Embodied agents admit no such shortcut, since they require d

safetyarxiv-cs-cv
28 Jul 2026
Model Releases

Do LLMs Know Their Vulnerable Scenarios?

DGX agent

arXiv:2607.23496v1 Announce Type: new Abstract: Safety-aligned large language models are trained to refuse harmful requests, yet embedding the same requests in particular scenarios can bypass their sa

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

DocHRL: A Hierarchical Reinforcement Learning Framework for Cost-Optimised Document Classification

DGX agent

arXiv:2607.22644v1 Announce Type: new Abstract: Real-world document classification pipelines typically apply the same sequence of models to every incoming document, regardless of its complexity or typ

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

DriveDNA: A Large-Scale Multimodal Naturalistic Driving Dataset and Benchmark for Driving Style Identification

DGX agent

arXiv:2607.23822v1 Announce Type: new Abstract: Driving style captures stable, driver-specific patterns in how a vehicle is driven. In naturalistic data, however, this signal is hard to isolate becaus

model-releasesarxiv-cs-lg
28 Jul 2026
Model Releases

FilmBench: A Film-Grade Benchmark for Cinematic Video Generation

DGX agent

arXiv:2607.24241v1 Announce Type: cross Abstract: Progress in video generation keeps narrowing the visual gap between AI-generated and professionally produced footage, yet most benchmarks still draw p

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

I built a tool to actually test which weights matter before quantizing, instead of guessing (Qwen3.6-27B, 3 builds: Bedrock/Tightrope/Gambit)

DGX agent

Most quantization works like this: pick a bit depth, apply it everywhere, maybe let imatrix take a rough guess at what matters, ship it. Most don't check which specific weight groups can take a hit an

model-releasesr-localllama
28 Jul 2026
Model Releases

If I only went off X posts, I'd think Ramp was an AI lab

DGX agent

If I only went off X posts, I'd think Ramp was an AI lab We’re open-sourcing PorTAL, our framework for shared task representations and cross model LoRA adaptation. It now spans from hybrid attention m

model-releasesjerry-liu--x
28 Jul 2026
Research

LA-RL: Label-Aware Self-Reflection for Reinforcement Learning in Information Extraction

DGX agent

arXiv:2607.23420v1 Announce Type: new Abstract: Large language models show strong promise for information extraction (IE), but existing reflection-based correction methods are often misaligned with st

researcharxiv-cs-cl
28 Jul 2026
Model Releases

Language Shapes Instruction Hierarchy Compliance in Multilingual LLMs

DGX agent

arXiv:2607.23545v1 Announce Type: new Abstract: Instruction hierarchy (IH) requires models to prioritize instructions by source, ensuring that higher-priority instructions override lower-priority ones

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

Mixture-of-Thought-Tokens: Unifying Perception and Reasoning for Free-form Multimodal Grounding

DGX agent

arXiv:2607.24407v1 Announce Type: new Abstract: Multimodal Large Language Models have made great progress in grounding tasks, yet existing methods still struggle to unify precise localization and comp

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

Multi-Modal Scene Graph with Kolmogorov-Arnold Experts for Audio-Visual Question Answering

DGX agent

arXiv:2511.23304v2 Announce Type: replace Abstract: In this paper, we propose a novel Multi-Modal Scene Graph with Kolmogorov-Arnold Expert Network for Audio-Visual Question Answering (SHRIKE). The ta

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Neuromorphic Object Detection: An In-Depth Study and Future Directions

DGX agent

arXiv:2607.23576v1 Announce Type: new Abstract: Conventional frame-based cameras face significant challenges in detecting objects under high-speed motion blur or in low-light environments. Neuromorphi

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

No Optimal Language Set Exists for Multilingual Instruction Tuning: Insights from a Linguistically-Informed Study

DGX agent

arXiv:2410.07809v2 Announce Type: replace Abstract: Multilingual instruction tuning (MIT) is challenged by the curse of multilinguality, data scarcity, and high computational cost. A natural hypothesi

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

On the Impossibility of Unbiased and Length-Invariant Policy Optimization with Outcome Rewards

DGX agent

arXiv:2607.23364v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) is the dominant reinforcement learning algorithm for training reasoning capabilities in large language models,

model-releasesarxiv-cs-lg
28 Jul 2026
Model Releases

OpenAIs HealthBench in Action: Evaluating an LLM-Based Medical Assistant on Realistic Clinical Queries

DGX agent

arXiv:2509.02594v3 Announce Type: replace-cross Abstract: Evaluating large language models (LLMs) on their ability to generate high-quality, accurate, situationally aware answers to clinical questions

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Parameter-Efficient Adaptation of SAM3 for Prompt-Driven Surgical Concept Segmentation

DGX agent

arXiv:2607.23694v1 Announce Type: new Abstract: Efficient surgical segmentation empowers clinical diagnosis, intraoperative monitoring, and downstream robotic pipelines for reconstruction and simulati

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

ParBench: A Benchmark for Reliable Evaluation of LLM Parallel Code Translation

DGX agent

arXiv:2607.22588v1 Announce Type: new Abstract: Modern compute-intensive software must migrate across a changing ecosystem of accelerators, programming APIs, compiler stacks, and portability layers, i

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Phenology-based learning framework for yield estimation and harvest forecasting of raspberry fruits

DGX agent

arXiv:2411.00967v2 Announce Type: replace Abstract: The future of agriculture is intertwined with automation. Accurate fruit detection, yield estimation, and harvest time prediction are crucial for ef

model-releasesarxiv-cs-cv
28 Jul 2026
Local Ai

Poison to Detect: Detection of Targeted Overfitting in Federated Learning

DGX agent

arXiv:2509.11974v3 Announce Type: replace-cross Abstract: Federated Learning (FL) enables collaborative model training among clients without centralising data, making it a widely adopted privacy-enhan

local-aiarxiv-cs-ai
28 Jul 2026
Model Releases

Poster: Rethinking Security in LLM Code Generation through Real-World Risk Scenarios

DGX agent

arXiv:2607.23088v1 Announce Type: cross Abstract: Large Language Models (LLMs) are widely used for code generation, yet their security behavior in realistic development workflows remains underexplored

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Random Forest-Based Prediction of Bone Volume Fraction and Fracture Position from S-Parameters

DGX agent

arXiv:2607.23563v1 Announce Type: new Abstract: In this paper, we propose a method for predicting bone volume fraction (BVF) and fracture position by constructing a random forest model based on multic

model-releasesarxiv-cs-lg
28 Jul 2026
Research

Same Question, Different Answers: Evaluating LLM Reliability Beyond Accuracy

DGX agent

arXiv:2607.22554v1 Announce Type: new Abstract: Large language models (LLMs) often achieve strong accuracy on benchmarks, yet it remains unclear how reliably they apply this knowledge when the same qu

researcharxiv-cs-ai
28 Jul 2026
Model Releases

Semalith v1.4: A Calibrated 184M Safety Classifier Achieving State-of-the-Art Prompt-Injection Detection at 44x Fewer Parameters than Llama-Guard-3-8B

DGX agent

arXiv:2607.22545v1 Announce Type: cross Abstract: Deploying large language models in financial-services and agentic settings requires safety classifiers that simultaneously handle prompt injection, re

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Source-Free Controlled Adaptation of Teachers for Continual Test-Time Adaptation

DGX agent

arXiv:2607.23735v1 Announce Type: cross Abstract: In many real-world scenarios, encountering continual shifts in domain during inference is very common. Consequently, continual test-time adaptation (C

model-releasesarxiv-cs-cv
28 Jul 2026
Hardware

Spatial-IQ: Deconstructing Spatial Intelligence via Hierarchical Capability Tests

DGX agent

arXiv:2607.22864v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) excel at visual interpretation but fail on spatial reasoning tasks that humans solve reliably. Existing bench

hardwarearxiv-cs-ai
28 Jul 2026
Model Releases

TextRich: A Multi-Domain Benchmark for Detecting AI-Generated Text-Rich Images from GPT-Image-2

DGX agent

arXiv:2606.19259v2 Announce Type: replace-cross Abstract: Text-rich images often contain privacy-sensitive, transactional, or decision-relevant information. As recent multimodal image generation model

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

TokenMem: Faithful Knowledge Injection for Frozen LLMs

DGX agent

arXiv:2607.22625v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) enhances large language models (LLMs) with external knowledge, but suffers from knowledge conflicts: when retrieved

model-releasesarxiv-cs-ai
28 Jul 2026
Research

VLASH: Real-Time VLAs via Future-State-Aware Asynchronous Inference

DGX agent

arXiv:2512.01031v2 Announce Type: replace-cross Abstract: Vision-Language-Action models (VLAs) are becoming increasingly capable across diverse robotic tasks. However, these models are typically deplo

researcharxiv-cs-ai
28 Jul 2026
Model Releases

Weighted Low-Rank Matrix Approximation: Acceleration and Applications

DGX agent

arXiv:2109.11057v2 Announce Type: replace-cross Abstract: Weighted low-rank matrix approximation (WLRMA) generalizes classical low-rank approximation and matrix completion by allowing arbitrary elemen

model-releasesarxiv-cs-lg
28 Jul 2026
Model Releases

XGRVFL-MV: Residual-Coupled Graph-Embedded Multi-View Random Vector Functional Link Network with FleXi Guardian Loss

DGX agent

arXiv:2607.23149v1 Announce Type: new Abstract: Random Vector Functional Link (RVFL) networks provide an efficient randomized learning framework for classification. Existing multi-view RVFL methods ut

model-releasesarxiv-cs-lg
28 Jul 2026
Model Releases

Data Quality over Capacity: Internalizing Documents into LoRA Adapters for Closed-Book QA

DGX agent

arXiv:2607.21861v1 Announce Type: new Abstract: We study baking documents directly into the weights of a 4-bit Gemma-4-e4b model via LoRA, so a system can answer questions about a corpus closed-book:

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

Do emulated quantum circuits change what CNNs look at? Performance and explainability comparison in medical image classification

DGX agent

arXiv:2607.21186v1 Announce Type: cross Abstract: Numerous studies have analyzed the use of hybrid quantum-classical convolutional neural networks as a promising alternative to classical deep learning

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

Filling Before Advancing: Capability-Gap-Driven Post-Training for Scenario-Specialized Remote Sensing MLLMs

DGX agent

arXiv:2607.22205v1 Announce Type: new Abstract: Remote sensing multimodal large language models (RS-MLLMs) have improved general aerial-image understanding. However, Earth observation applications req

model-releasesarxiv-cs-cv
27 Jul 2026
Model Releases

Happy to have @FireworksAI_HQ as our day0 launch partner and bring Kimi K3 to more developers. With Fireworks, you can now deploy and fine-t…

DGX agent

Happy to have @FireworksAI_HQ as our day0 launch partner and bring Kimi K3 to more developers. With Fireworks, you can now deploy and fine-tune the 2.8T Kimi K3 model with just a few clicks! Kimi K3 i

model-releaseskimi-moonshot--x
27 Jul 2026
Model Releases

Interpretable EEG biomarkers with bag-of-waves: Spatial and temporal waveform dictionaries for low-data regimes

DGX agent

arXiv:2607.22508v1 Announce Type: new Abstract: Electroencephalography (EEG) is widely used to diagnose neurological conditions, but its analysis usually relies on either predefined spectral features

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

Learning What Matters: Supervising Sparse Attention Routing with Causal Evidence Sets

DGX agent

arXiv:2607.21692v1 Announce Type: cross Abstract: Sparse attention reduces the cost of long contexts by allowing each query to read only selected parts of the input. These selectors are often trained

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

Modernizing the skies: NOAA and Google Cloud collaborate to advance weather forecasting

DGX agent

The National Oceanic and Atmospheric Administration (NOAA) is embarking on a transformative journey to redefine how we understand and predict patterns in the Earth’s atmosphere that affect the weather

model-releasesgoogle-cloud-ai
27 Jul 2026
Research

Multi-Horizon Consistency as Geometry: When Latent Dynamics Contract, and When They Do Not

DGX agent

arXiv:2607.21645v1 Announce Type: new Abstract: Multi-horizon latent consistency is a common training knob in video predictors and world models, but practitioners rarely know what it does to transitio

researcharxiv-cs-lg
27 Jul 2026
Model Releases

Procedural Knowledge Is Not Low-Rank: Why LoRA Fails to Internalize Multi-Step Procedures

DGX agent

arXiv:2607.21612v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning methods like LoRA have become the default for adapting large language models, succeeding across instruction following,

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

SceneActBench: Can Agents Act on the 3D Scenes They See?

DGX agent

arXiv:2607.22393v1 Announce Type: cross Abstract: Vision-language model (VLM) agents increasingly use tools to act on 3D scenes rather than only describe them. Existing 3D benchmarks score textual res

model-releasesarxiv-cs-cv
27 Jul 2026
Model Releases

Spatially-Enhanced Temporal Fusion Transformer: Interpretable Multi-Output Prediction for Parametric Dynamical Systems with Time-Varying Inputs

DGX agent

arXiv:2505.00473v2 Announce Type: replace Abstract: We explore the promising performance of a transformer model in predicting outputs of parametric dynamical systems with external time-varying input s

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

Variational Low-rank Tensor Decomposition for Multisubject Spatiotemporal Data Analysis

DGX agent

arXiv:2607.22262v1 Announce Type: cross Abstract: Modeling shared and subject-specific structure in multisubject spatiotemporal data remains challenging, particularly in neuroimaging, where both spati

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

Ha! It did it: 'We introduce BenchBenchBenchBenchBench (BBBBB), an executable benchmark of AI-authored conformance suites for benchmark-eval…

DGX agent

Ha! It did it: 'We introduce BenchBenchBenchBenchBench (BBBBB), an executable benchmark of AI-authored conformance suites for benchmark-evaluation metrics' I really thought it would treat 'now do benc

model-releasesethan-mollick--x
25 Jul 2026
Model Releases

Ollama Cloud Quota Benchmark

DGX agent

Recently I bought an Ollama Cloud sub and accidently spent my whole 5h quota upon using DeepSeek V4 Pro... but why? isnt it supposed to be a cheap model? Youd think there would be a correlation betwee

model-releasesr-ollama
25 Jul 2026
Model Releases

Quoting Boris Cherny

DGX agent

More than any of these eval scores, what is most exciting to me is something else: Opus 5 is our least prompt injectable model yet. It is a bit buried in the system card, but across PI evals and red t

model-releasessimon-willison
25 Jul 2026
Model Releases

Anthropic releases Opus 5 with ‘close’ to Fable 5’s capabilities

DGX agent

Weeks after Anthropic's latest toe-to-toe with the US government, and days after an OpenAI security incident that dominated tech industry discussions, Anthropic on Thursday released its newest model,

model-releasesthe-verge-ai
24 Jul 2026
Safety

excellent

DGX agent

excellent For my first post, I’m sharing a letter @NVIDIA signed on why open models matter. AI will transform every industry, power every company, and be built by every country. Open models strengthen

safetyelon-musk--x
24 Jul 2026
← Previous
1…385386387388389…1326
Next →