AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

85,202Total entries
1Added by human
85,201Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,047 results
17 Aug 2026

Revisiting Shape and Texture Reliance with Category-Separability-Calibrated Suppression

Model ReleasesDGX agent

arXiv:2607.16298v2 Announce Type: replace Abstract: Feature-suppression evaluations infer model reliance on shape or texture from the accuracy loss caused by attenuating each type of information. Such

When Denoising Hurts: Rethinking the Terminal Step of Diffusion Time Series Forecasters -- Extended Version

ApplicationsDGX agent

arXiv:2608.14067v1 Announce Type: new Abstract: Diffusion models offer a natural way to model uncertainty in time series forecasting, yet their iterative sampling process is often treated as a uniform

16 Aug 2026

Qwen3.8 27B Q2 vs Q3 vs Qwen3.6 35B-A3B MoE on 12GB VRAM

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

Did a quick local test because I wanted to see what is actually usable on my 12GB laptop GPU. I tested the newer Qwen3.8 27B dense files at Q2 and Q3, then compared them against Qwen3.6 35B-A3B MoE. H

15 Aug 2026

llama.cpp Windows Manager

Model ReleasesDGX agent

https://github.com/alekk89/llama-cpp-windows-manager It has been three months since I first shared my personal solution for running llama.cpp on Windows, and the project has evolved considerably since

14 Aug 2026

BavGround: A Benchmark for Regional Cultural Grounding and Dialect Competence in Bavarian

Model ReleasesDGX agent

arXiv:2608.12894v1 Announce Type: new Abstract: Cultural evaluation of large language models (LLMs) often focuses on high-resource standard languages, leaving regional culture and dialect communities

CangjieBench: Benchmarking LLMs on a Low-Resource General-Purpose Programming Language

Model ReleasesDGX agent

arXiv:2603.14501v2 Announce Type: replace-cross Abstract: Large Language Models excel in high-resource programming languages but struggle with low-resource ones. Existing research related to low-resou

CityRiSE: Reasoning Urban Socio-Economic Status in Large Vision-Language Models via Reinforcement Learning

ResearchDGX agent

arXiv:2510.22282v2 Announce Type: replace-cross Abstract: Urban socio-economic sensing plays a vital role in advancing global sustainable development goals. With the advent of Large Vision-Language Mo

CW-BASS v2: Saturation-Aware Pseudo-Label Selection for Semi-Supervised Segmentation under Foundation-Model Teachers

SafetyDGX agent

arXiv:2608.12773v1 Announce Type: new Abstract: Semi-supervised semantic segmentation has long turned on one question, which pseudo-labels to trust, and a generation of selection rules, dynamic thresh

DFM Mimir v1: An Open HRM Delivering Frontier Performance at 1B Parameters Using Only Permissible Post-Training Data

Model ReleasesDGX agent

arXiv:2608.13517v1 Announce Type: cross Abstract: Current large language model development relies on massive, often non-permissible datasets, creating a high barrier for researchers committed to open-

DMDIntel: Interpreting Large Language Models via Dynamic Mode Decomposition

ResearchDGX agent

arXiv:2608.13048v1 Announce Type: new Abstract: In this work, we introduce DMDIntel which uses dynamic mode decomposition (DMD) to make the predictions made by LLMs in a classification task interpreta

EgoMonth: A Month-Level Egocentric Video Benchmark for Long-Term Spatiotemporal Memory

Model ReleasesDGX agent

arXiv:2608.13113v1 Announce Type: cross Abstract: Recent advances in Multimodal Large Language Models (MLLMs) have led to substantial progress in video understanding, accompanied by a growing number o

Evaluation Resolution Confounds Learning-Rule Comparisons in Model-Brain RSA of Early Visual Cortex

SafetyDGX agent

arXiv:2608.12408v1 Announce Type: cross Abstract: Representational similarity analysis (RSA) is increasingly used to ask which learning rules give convolutional networks brain-like representations. Be

Falsehood and Impossibility Are Different Directions in an AI's Representation of Language

Model ReleasesDGX agent

arXiv:2608.12852v1 Announce Type: cross Abstract: Language can describe states of affairs that are false and states of affairs that could not be the case at all. Whether an AI model internally disting

GS^{2}CI: Robust Gaussian Splatting For Snapshot Compressive Imaging via Large Vision Model Priors

ResearchDGX agent

arXiv:2608.13502v1 Announce Type: new Abstract: Snapshot Compressive Imaging (SCI) offers an efficient solution for high-speed video acquisition and, under exposure-time camera--scene relative motion,

I made a harness for Ollama

Local AiDGX agent

It'll use whatever model you have installed on ollama. Pro+/Cloud accounts are supported and you can ensemble multiple models and compound their results into one response or individually. Looking for

Qwen 3.8 27B - Aquarium Burst Sample Test

Model ReleasesDGX agent

Tested this prompt on the full version (BF16). Though this was a single prompt, I executed using vscode GH copilot extension on agent (allow all) mode and let it do its thing. So there were 54 model t

Refusing Intent, Not Form: Wrapper-Based Intent-Group Supervision for LLM Safety

Model ReleasesDGX agent

arXiv:2608.13304v1 Announce Type: new Abstract: Safety tuning can improve harmful refusal, but models may learn surface-form shortcuts: wrapped harmful prompts bypass safety, while similarly wrapped b

Which LLM Is Your Ideal Companion? Evaluating Emotional Companion Capabilities of LLMs Based on Adult Attachment Theory

Model ReleasesDGX agent

arXiv:2608.13168v1 Announce Type: new Abstract: As large language models (LLMs) are increasingly applied for emotional companionship, evaluating their behavior and capabilities in intimate relationshi

13 Aug 2026

Beyond Parameter Space: NTK-Guided Personalized Aggregation for Robust Federated Learning

Model ReleasesDGX agent

arXiv:2608.12108v1 Announce Type: new Abstract: Federated learning (FL) enables collaborative model training across distributed clients while keeping data local. A central challenge is determining whi

COGENT: Counterfactual Gaussian Explanations for Volumetric Medical Images

Model ReleasesDGX agent

arXiv:2608.11422v1 Announce Type: new Abstract: Explainability is essential for deploying deep learning models in high-stakes medical applications. Existing explainability methods for volumetric imagi

Decoupled Quadratic Kalman Filter for Elliptical Extended Object Tracking with Log-normal Axis Modeling

ResearchDGX agent

arXiv:2512.14426v2 Announce Type: replace-cross Abstract: Extended object tracking involves estimating both the physical extent and kinematic parameters of a target object, where typically multiple me

EXPERIMENT: Qwen3.8-2.4T-A95B running locally on an RTX 5090 + RTX 5060 Ti at ~0.80 tok/s

Model ReleasesDGX agent

I managed to get Qwen3.8-2.4T-A95B running locally with llama.cpp on mu PC just for fun, cause why not. I was using the Unsloth Qwen3.8-2.4T-A95B-UD-Q1_0 GGUF quantization. The full GGUF is about 397

LabelFusion-TS: Fusing Large Language Models, Transformer Encoders, and Financial Time Series for Monetary-Policy Stance Classification

SafetyDGX agent

arXiv:2608.11753v1 Announce Type: new Abstract: Financial text is produced and interpreted within a market environment, yet financial text classifiers almost always receive text alone. We study whethe

Massive Activations in Hybrid Linear Attention Large Language Models: Pre-Attention Spikes and Inter-Spike Plateaus

ResearchDGX agent

arXiv:2608.12149v1 Announce Type: new Abstract: We present the first systematic study of Massive activations (MAs) in layer-interleaved HLA LLMs and uncover two architecture-aligned morphologies: MAs

ScreenShot: A Foundation Model for Few-Shot Combination Drug Screening

ResearchDGX agent

arXiv:2608.12219v1 Announce Type: new Abstract: Treating patients with combinations of drugs reduces the risk of resistance to any individual drug. Finding effective combinations is difficult because

Self-Evolving Code-with-Image Reasoning

Model ReleasesDGX agent

arXiv:2608.11292v1 Announce Type: new Abstract: Multimodal models increasingly reach for tools when solving visual tasks (crop, zoom, rotate, brighten), a paradigm known as thinking-with-images. The c

TangPoetryBench: A Multi-Dimensional Benchmark and Rubric-Conditioned Evaluator for Poetry-to-Image Generation

Model ReleasesDGX agent

arXiv:2608.11452v1 Announce Type: cross Abstract: Text-to-image (T2I) models are increasingly asked to illustrate literary and cultural content, yet we cannot measure how well an image renders the mea

The top 10% of enterprises use plugins twice as often and skills six times as often as typical firms. These frontier firms are not ahead by …

Model ReleasesDGX agent

According to a Twitter post by OpenAI on August 13, 2026, the top 10% of enterprises adopt AI plugins twice as frequently and leverage skills—such as model‐based knowledge and reasoning—six times more

Towards foundation-style models for energy-frontier heterogeneous neutrino detectors via self-supervised pre-training

ResearchDGX agent

arXiv:2604.07037v2 Announce Type: replace-cross Abstract: Accelerator-based neutrino physics is entering an energy-frontier regime in which interactions reach the TeV scale and produce exceptionally d

Towards Model-based Run-time Cybersecurity: On Control-Flow Anomaly Detection, Attack Identification, and Hardware Monitoring

ResearchDGX agent

arXiv:2608.11802v1 Announce Type: cross Abstract: Methods to increase the resilience of systems to cyber-attacks become increasingly important. Control-flow monitoring provides a principled basis to e

When the API Speaks the Wrong Language: Revisiting Post-Training for Multilingual Tool Use

Model ReleasesDGX agent

arXiv:2608.11715v1 Announce Type: cross Abstract: The reliability of Large Language Models (LLMs) for API calling degrades in multilingual settings. A common failure occurs when a model selects the co

Zero-OVCD: Bridging Training-Free Foundation Models and Pseudo-Label Learning for Open-Vocabulary Change Detection

ResearchDGX agent

arXiv:2608.11663v1 Announce Type: new Abstract: Open-vocabulary change detection (OVCD) enables the identification of user-specified land-cover changes in bitemporal remote sensing images, but existin

12 Aug 2026

Actionable Hallucination Detection: Translating Latent Uncertainty into Agentic Critique

Model ReleasesDGX agent

arXiv:2608.10430v1 Announce Type: cross Abstract: Large Language Models (LLMs) deployed as AI agents frequently exhibit user specification-grounding failures, executing hallucinated, undesired actions

Actions Speak Louder than Words: Measuring Cross-Lingual Policy Retention in Tool-Using Agents

SafetyDGX agent

arXiv:2608.11110v1 Announce Type: new Abstract: When a tool-using agent is given the same task in a different language, does it still take the same steps? Multilingual evaluation rarely asks: it compa

Cracks in the Foundation: Seemingly Minor Architectural Choices Impact Long Context Extension

Model ReleasesDGX agent

arXiv:2608.10296v1 Announce Type: new Abstract: One might imagine that architectural variations within the dense transformer paradigm have a limited effect on accuracy. However, we demonstrate that th

Data Attribution of Emergent Misalignment with Persona Features

SafetyDGX agent

arXiv:2608.11025v1 Announce Type: new Abstract: Emergent misalignment (EM) is the phenomenon where fine-tuning a language model on a narrow task leads to harmful behavior in unrelated domains. A leadi

Evidence-Grounded Trustworthy Multimodal Reasoning and Evaluation Benchmark in Complex Urban Scenes

Model ReleasesDGX agent

arXiv:2608.10954v1 Announce Type: cross Abstract: While Multimodal Large Language Models (MLLMs) demonstrate impressive performance in benign scenarios, their cognitive reliability deteriorates signif

Interesting research suggests caution in determining which AI company is winning by looking at any one source.. OpenRouter seems to show ope…

Model ReleasesDGX agent

Interesting research suggests caution in determining which AI company is winning by looking at any one source.. OpenRouter seems to show open weights winning over time, but work submitted to Pangram i

Is This Your Final Answer? Cross-Contextual Consistency as a Measure of LLM Credibility

Model ReleasesDGX agent

arXiv:2608.10315v1 Announce Type: cross Abstract: Large language models (LLMs) are powerful black-box systems, making it difficult to discern whether their answers reflect stable internal beliefs or s

JEPA-WAM: Stage-Level Joint-Embedding Prediction for World-Action Models in Robot Manipulation

Local AiDGX agent

arXiv:2608.10780v1 Announce Type: new Abstract: Generalist robot policies aim to map multimodal observations and linguistic task instructions to actions across diverse tasks. However, existing methods

Longitudinal 3D Foundation Modeling for Neoadjuvant Breast Cancer Response Prediction from Serial DCE-MRI

ResearchDGX agent

arXiv:2608.09991v1 Announce Type: cross Abstract: Pathologic complete response (pCR) is an important endpoint in neoadjuvant chemotherapy (NAC) for breast cancer, and predicting pCR from imaging durin

MemSpec: Memory-Aware Runtime for Adaptive Draft Scheduling in Speculative Decoding on Edge Devices

ResearchDGX agent

arXiv:2608.10362v1 Announce Type: cross Abstract: Speculative decoding accelerates autoregressive large language model (LLM) inference by using a lightweight draft model to speculate multiple tokens,

myMediWhisper: Construction of Burmese Medical Speech Corpus and Whisper Fine-Tuning for Clinical Dialogue ASR

Model ReleasesDGX agent

arXiv:2608.11036v1 Announce Type: new Abstract: Although Whisper models benefit from large-scale multilingual pre-training, their performance on Burmese medical speech remains limited. This work prese

Nonlinear Model Predictive Control via Sequential Convex Programming for Drone-to-Drone Docking

AgentsDGX agent

arXiv:2608.10542v1 Announce Type: new Abstract: Autonomous mid-air docking of multi-rotor vehicles under disturbance-driven target motion poses a constrained non-linear trajectory optimization challen

Robust and Secure Code Watermarking for Large Language Models via ML/Crypto Codesign

ResearchDGX agent

arXiv:2502.02068v3 Announce Type: replace-cross Abstract: This paper introduces RoSeMary, the first-of-its-kind ML/Crypto codesign watermarking framework that regulates LLM-generated code to avoid int

Status Association Does Not Reliably Predict Decision Leakage

SafetyDGX agent

arXiv:2608.10089v1 Announce Type: cross Abstract: Bias evaluations often move too quickly from evidence that a model encodes a social association to claims that the same association will alter consequ

Stream Forcing: Constructing Unified Training Trajectory for Robust Streaming Video Generation

Model ReleasesDGX agent

arXiv:2608.10439v1 Announce Type: new Abstract: Streaming video generation holds strong potential for world modeling, where future frames must be inferred online sequentially to form a continuous vide

Toward a Theory of Value in AI Alignment

SafetyDGX agent

arXiv:2608.10327v1 Announce Type: new Abstract: Can AI systems be aligned to human values? The popularization of large language models (LLMs) and multi-modal foundation models has seen a rise in harms

v0.32.10-rc0: nn: speed up prefill on double-scale nvfp4 models

Local AiDGX agent

ModelOpt checkpoints apply a float32 global scale to every projection output on top of the per-group quantization scales. Running the multiply and the cast back to the activation dtype as separate eag

With the smartest person I know, @HarshSensei, we're building evsys-sdk for the community, this open-source repository allows anyone to buil…

Model ReleasesDGX agent

With the smartest person I know, @HarshSensei, we're building evsys-sdk for the community, this open-source repository allows anyone to build their own continual learning system with first-class suppo

Workflow Cards: Structured Summaries of Workflow Executions Using Provenance Data

Model ReleasesDGX agent

arXiv:2608.11022v1 Announce Type: cross Abstract: Model Cards and Data Cards have demonstrated the value of structured, human-readable documentation for machine learning artifacts, capturing their con

11 Aug 2026

Calib3R: Hand-Eye Calibration and 3D Metric-Scaled Scene Reconstruction with 3D Foundation Models

ResearchDGX agent

arXiv:2509.08813v2 Announce Type: replace Abstract: Robots often rely on RGB images for tasks like manipulation. However, reliable interaction typically requires a 3D scene representation that is metr

Can Webcam Gaze Constrain Mesa-Objectives in Driving Models? An Instrument Precision Analysis

AgentsDGX agent

arXiv:2608.08947v1 Announce Type: new Abstract: Current hazard detection systems in autonomous driving may develop mesa objectives, learned internal goals that achieve high training performance throug

Classical SU(2) Models Match or Exceed Shallow Variational Quantum Circuits on Vision Benchmarks

Local AiDGX agent

arXiv:2608.07822v1 Announce Type: cross Abstract: Quaternion-valued neural networks and variational quantum circuits (VQCs) both derive local transformations from SU(2) geometry, yet their performance

CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Continuous Latents

ResearchDGX agent

arXiv:2608.08638v1 Announce Type: cross Abstract: Zero-shot text-to-speech (TTS) now supports interactive assistants, personalized media, and accessibility tools. All TTS systems require faithful ling

Decoding Phenotypes: A Framework for Fusing Genomic Language Models and Neuroimaging

ResearchDGX agent

arXiv:2608.08926v1 Announce Type: new Abstract: Neuroimaging and genetic testing are two important clinical references for nervous system diseases, offering complementary diagnostic information. Howev

DINO-3DRA: Leveraging 2D Foundation Model Semantics for 3D Cerebral Aneurysm Segmentation

ResearchDGX agent

arXiv:2608.07767v1 Announce Type: new Abstract: Accurate aneurysm segmentation in 3D rotational angiography (3DRA) is hindered by extreme class imbalance, morphological similarity to vessels, and abse

DS@GT ARC at Touche: Large Language Models for Retrieval-Augmented Debate

ResearchDGX agent

arXiv:2608.08143v1 Announce Type: cross Abstract: We extend the DS@GT ARC working-note submission to the Touche 2025 Retrieval-Augmented Debate task. The task has two subtasks: generating the next utt

Emotion2Skill: Model-Internal Emotion Signals for Adaptive Skill Selection and Evolution

AgentsDGX agent

arXiv:2608.09248v1 Announce Type: new Abstract: Skill-based LLM agents select reusable procedures from an external library to solve complex tasks, yet their routing decisions rely entirely on text-lev

Exclusive: ZeroDrift applies small language model to prevent AI-generated compliance violations

IndustryDGX agent

“This investment is guaranteed to return 12% annually.” A claim like that in an email from an investment adviser is a regulatory disaster. Regulations prohibit financial firms from promising returns a

← Previous
1…243244245246247…1018
Next →