AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlog
86,428Total entries
1Added by human
86,427Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,016 results
Safety

Bias-Constrained Diffusion Schedules for PDE Emulations: Reconstruction Error Minimization and Efficient Unrolled Training

DGX agent

arXiv:2604.08357v2 Announce Type: replace Abstract: Conditional Diffusion Models are powerful surrogates for emulating complex spatiotemporal dynamics, yet they often fail to match the accuracy of det

safetyarxiv-cs-lg
13 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

On-the-Fly Adaptation to Quantization: Configuration-Aware LoRA for Efficient Fine-Tuning of Quantized LLMs

DGX agent

arXiv:2509.25214v3 Announce Type: replace-cross Abstract: As increasingly large pre-trained models are released, deploying them on edge devices for privacy-preserving applications requires effective c

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

SiMing-Bench: Evaluating Procedural Correctness from Continuous Interactions in Clinical Skill Videos

DGX agent

arXiv:2604.09037v1 Announce Type: cross Abstract: Current video benchmarks for multimodal large language models (MLLMs) focus on event recognition, temporal ordering, and long-context recall, but over

model-releasesarxiv-cs-cl
13 Apr 2026
Applications

Task-Distributionally Robust Data-Free Meta-Learning

DGX agent

arXiv:2311.14756v2 Announce Type: replace-cross Abstract: Data-Free Meta-Learning (DFML) aims to enable efficient learning of unseen few-shot tasks, by meta-learning from multiple pre-trained models w

applicationsarxiv-cs-ai
13 Apr 2026
Model Releases

After seeing that Claude Mythos marketing turned out to be, as expected, a scam, I wanted to make a master list of tricks being used to mark…

DGX agent

After seeing that Claude Mythos marketing turned out to be, as expected, a scam, I wanted to make a master list of tricks being used to market LLMs. The master list includes statements directly from l

model-releasesyann-lecun--x
12 Apr 2026
Model Releases

Action Without Interaction: Probing the Physical Foundations of Video LMMs via Contact-Release Detection

DGX agent

arXiv:2511.20162v2 Announce Type: replace Abstract: Large multi-modal models (LMMs) show increasing performance in realistic visual tasks for images and, more recently, for videos. For example, given

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Break Me If You Can: Self-Jailbreaking of Aligned LLMs via Lexical Insertion Prompting

DGX agent

arXiv:2601.02670v2 Announce Type: replace Abstract: We introduce self-jailbreaking, a threat model in which an aligned LLM guides its own compromise. Unlike most jailbreak techniques, which oft

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Calibration of a neural network ocean closure for improved mean state and variability

DGX agent

arXiv:2604.06398v1 Announce Type: cross Abstract: Global ocean models exhibit biases in the mean state and variability, particularly at coarse resolution, where mesoscale eddies are unresolved. To add

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

Diagnosing and Mitigating Sycophancy and Skepticism in LLM Causal Judgment

DGX agent

arXiv:2601.08258v3 Announce Type: replace Abstract: Large language models increasingly fail in a way that scalar accuracy cannot diagnose: they produce a sound reasoning trace and then abandon it unde

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Dynamic Context Evolution for Scalable Synthetic Data Generation

DGX agent

arXiv:2604.07147v1 Announce Type: cross Abstract: Large language models produce repetitive output when prompted independently across many batches, a phenomenon we term cross-batch mode collapse: the p

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

GLM-5.1 by @Zai_org is now #3 in Code Arena - surpassing Gemini 3.1 and GPT-5.4, and now on par with Claude Sonnet 4.6. The first frontier l…

DGX agent

GLM-5.1 by @Zai_org is now #3 in Code Arena - surpassing Gemini 3.1 and GPT-5.4, and now on par with Claude Sonnet 4.6. The first frontier level open model to break into the top 3. It’s a major +90 po

model-releaseszhipu-ai--x
10 Apr 2026
Model Releases

JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiency

DGX agent

arXiv:2604.03044v2 Announce Type: replace-cross Abstract: We introduce JoyAI-LLM Flash, an efficient Mixture-of-Experts (MoE) language model designed to redefine the trade-off between strong performan

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

LLM Spirals of Delusion: A Benchmarking Audit Study of AI Chatbot Interfaces

DGX agent

arXiv:2604.06188v1 Announce Type: cross Abstract: People increasingly hold sustained, open-ended conversations with large language models (LLMs). Public reports and early studies suggest that, in such

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

LoFT: Parameter-Efficient Fine-Tuning for Long-tailed Semi-Supervised Learning in Open-World Scenarios

DGX agent

arXiv:2509.09926v5 Announce Type: replace Abstract: Long-tailed semi-supervised learning (LTSSL) presents a formidable challenge where models must overcome the scarcity of tail samples while mitigatin

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

MinerU2.5-Pro: Pushing the Limits of Data-Centric Document Parsing at Scale

DGX agent

arXiv:2604.04771v2 Announce Type: replace-cross Abstract: Current document parsing methods advance primarily through model architecture innovation, while systematic engineering of training data remain

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

mlx @ollama!

DGX agent

Ollama released a preview version (0.19) on March 31, 2026, built on top of Apple's open-source MLX framework, enabling local LLMs to run significantly faster on Apple Silicon Macs by leveraging th...

model-releasesollama--x
10 Apr 2026
Model Releases

PLUME: Latent Reasoning Based Universal Multimodal Embedding

DGX agent

arXiv:2604.02073v2 Announce Type: replace Abstract: Universal multimodal embedding (UME) maps heterogeneous inputs into a shared retrieval space with a single model. Recent approaches improve UME by g

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

PSR: Scaling Multi-Subject Personalized Image Generation with Pairwise Subject-Consistency Rewards

DGX agent

arXiv:2512.01236v2 Announce Type: replace Abstract: Personalized generation models for a single subject have demonstrated remarkable effectiveness, highlighting their significant potential. However, w

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Reinforcement-Guided Synthetic Data Generation for Privacy-Sensitive Identity Recognition

DGX agent

arXiv:2604.07884v1 Announce Type: new Abstract: High-fidelity generative models are increasingly needed in privacy-sensitive scenarios, where access to data is severely restricted due to regulatory an

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

SALLIE: Safeguarding Against Latent Language & Image Exploits

DGX agent

arXiv:2604.06247v1 Announce Type: cross Abstract: Large Language Models (LLMs) and Vision-Language Models (VLMs) remain highly vulnerable to textual and visual jailbreaks, as well as prompt injections

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

T-Gated Adapter: A Lightweight Temporal Adapter for Vision-Language Medical Segmentation

DGX agent

arXiv:2604.08167v1 Announce Type: new Abstract: Medical image segmentation traditionally relies on fully supervised 3D architectures that demand a large amount of dense, voxel-level annotations from c

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

TEMPER: Testing Emotional Perturbation in Quantitative Reasoning

DGX agent

arXiv:2604.07801v1 Announce Type: new Abstract: Large language models are trained and evaluated on quantitative reasoning tasks written in clean, emotionally neutral language. However, real-world quer

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

What’s new with Google Cloud

DGX agent

Want to know the latest from Google Cloud? Find it here in one handy location. Check back regularly for our newest updates, announcements, resources, events, learning opportunities, and more. Tip: Not

model-releasesgoogle-cloud-ai
10 Apr 2026
Model Releases

Gemma 4 31B brings dense multimodal reasoning to Together AI. Try Now: http://www.together.ai/models/gemma-4-31b

DGX agent

Google's Gemma 4 31B is a dense multimodal model from Google DeepMind now available on Together AI's serverless infrastructure via the endpoint `google/gemma-4-31B-it`. It features a 256K context ...

model-releasestogether-ai--x
9 Apr 2026
Model Releases

Judging by my tl there is a growing gap in understanding of AI capability. The first issue I think is around recency and tier of use. I thin…

DGX agent

Judging by my tl there is a growing gap in understanding of AI capability. The first issue I think is around recency and tier of use. I think a lot of people tried the free tier of ChatGPT somewhere l

model-releaseskarpathy--x
9 Apr 2026
Agents

SWE-1.6 is lightning fast! Here's what 950 tok/s feels like - available in Windsurf today.

DGX agent

SWE-1.6 is lightning fast! Here's what 950 tok/s feels like - available in Windsurf today. Media We’re releasing SWE-1.6, our best model in both intelligence & model UX. SWE-1.6 matches our Preview mo

agentscognition-ai--x
8 Apr 2026
Agents

Or try SWE-1.6 in Windsurf today: https://windsurf.com/

DGX agent

Cognition released SWE-1.6, their latest software engineering model optimized for both intelligence and 'model UX,' now generally available in the Windsurf IDE. It is free for the next three month...

agentscognition-ai--x
7 Apr 2026
Safety

Advanced modelling and data analytics in aviation

DGX agent

arXiv:2608.14746v1 Announce Type: new Abstract: The aviation industry characterized by its stringent safety standards has seen a growing need for innovative approaches to enhance safety measures. Desp

safetyarxiv-cs-ai
18 Aug 2026
Hardware

AlloEgo-VLM: Disambiguating Allocentric and Egocentric Reference Frames in Vision-Language Models

DGX agent

arXiv:2608.15605v1 Announce Type: new Abstract: This study investigates the challenge of ambiguity faced by Vision-Language Models (VLMs) in understanding spatial semantics. Spatial cognition, shaped

hardwarearxiv-cs-cv
18 Aug 2026
Research

AnyTalk: Speech Animation for Arbitrary Characters Leveraging a Video Generation Model

DGX agent

arXiv:2608.16143v1 Announce Type: cross Abstract: We present AnyTalk, a novel method for generating 3D speech animations for arbitrary characters without requiring any animation data. While existing a

researcharxiv-cs-cv
18 Aug 2026
Applications

Arm-Aware Guided Dexterous Grasp Generation with Arm-Agnostic Grasp Models

DGX agent

arXiv:2608.16351v1 Announce Type: new Abstract: Dexterous grasp generation that considers arm-related constraints is crucial in real-world scenarios involving arm environment collision avoidance, work

applicationsarxiv-cs-ro
18 Aug 2026
Local Ai

b10481: CUDA: MMVQ nwarps=8 for bs=1 for dense models on DGX Spark (#26843)

DGX agent

CUDA: MMVQ nwarps=8 for bs=1 for dense models on DGX Spark Signed-off-by: ynankani ynankani@nvidia.com skip moe experts and allow others based on k geometry (allow only small idle tail) Signed-off-by:

local-aillama-cpp-releases
18 Aug 2026
Research

Doubly robust nearest neighbors in factor models

DGX agent

arXiv:2211.14297v4 Announce Type: replace-cross Abstract: We introduce and analyze an improved variant of nearest neighbors (NN) for estimation with missing data in latent factor models. We consider a

researcharxiv-cs-lg
18 Aug 2026
Agents

DriveCache: Action-Aware Caching for Driving World Model Inference

DGX agent

arXiv:2608.16354v1 Announce Type: new Abstract: Driving video generation models support autonomous-driving development by predicting controllable future scenes for simulation, planning evaluation, and

agentsarxiv-cs-ai
18 Aug 2026
Research

Dynamic Multi-Byte Prediction With Hierarchical Language Models

DGX agent

arXiv:2608.15454v1 Announce Type: new Abstract: Byte-level hierarchical language models (LMs) have recently emerged as a robust alternative to their popular counterparts that use subword tokenization.

researcharxiv-cs-ai
18 Aug 2026
Research

H-PAC Hand: Control-Oriented Modeling and Tendon-Elasticity Compensation for an Underactuated Robotic Hand

DGX agent

arXiv:2608.16712v1 Announce Type: new Abstract: Underactuated tendon-driven hands offer compact actuation and passive compliance, but tendon elongation under restoring-spring loading introduces config

researcharxiv-cs-ro
18 Aug 2026
Tutorials

HOIMask: Towards Generative Masked Modeling for Human Object Interaction Generation

DGX agent

arXiv:2608.15141v1 Announce Type: new Abstract: Diffusion-based methods have dominated the HOI generation, as they enable critical contact fusions or signals to guide the diffusion process. However, t

tutorialsarxiv-cs-cv
18 Aug 2026
Safety

I-Perceive: A Foundation Model for Active Perception with Language Instructions

DGX agent

arXiv:2603.00600v2 Announce Type: replace Abstract: Active perception - the ability of a robot to proactively select viewpoints to acquire task-relevant information - is essential for robust operation

safetyarxiv-cs-ro
18 Aug 2026
Safety

Imagining Recovery: Inference-Time Counterfactual Realignment for Vision-Language-Action Models

DGX agent

arXiv:2608.14822v1 Announce Type: cross Abstract: Vision-language-action (VLA) models have improved the flexibility and generality of robotic manipulation, yet they remain fragile to online disruption

safetyarxiv-cs-cv
18 Aug 2026
Applications

Macroeconomic Forecasting with Large Language Models

DGX agent

arXiv:2407.00890v5 Announce Type: replace-cross Abstract: This paper presents a comparative analysis evaluating the accuracy of Large Language Models (LLMs) against traditional macro time series forec

applicationsarxiv-cs-cl
18 Aug 2026
Tutorials

NebulaVLA: A Dual-Frequency Vision-Language-Action Model With Guide Action for Robotic Manipulation

DGX agent

arXiv:2608.16503v1 Announce Type: cross Abstract: Real-world deployment of Vision-Language-Action (VLA) models is often bottlenecked by efficiency-performance trade-offs, cross-embodiment generalizati

tutorialsarxiv-cs-ai
18 Aug 2026
Research

Paired Exact-Reset Evaluation of a Prediction-Derived Medium-to-Full World-Model Cascade

DGX agent

arXiv:2608.14650v1 Announce Type: new Abstract: Existing adaptive-inference and world-action-model systems use cheap-stage outputs or predicted futures to allocate additional computation. We study a n

researcharxiv-cs-lg
18 Aug 2026
Applications

PhyxMamba: Chaotic System Reconstruction from Short Context Observations with Generative State-Space Models

DGX agent

arXiv:2505.23863v3 Announce Type: replace-cross Abstract: Understanding chaotic dynamics is a fundamental problem across scientific disciplines, including climate science, neuroscience, and fluid dyna

applicationsarxiv-cs-ai
18 Aug 2026
Agents

RingMo-Agent: A Unified Remote Sensing Foundation Model for Multi-Platform and Multi-Modal Reasoning

DGX agent

arXiv:2507.20776v3 Announce Type: replace Abstract: Remote sensing (RS) images from multiple modalities and platforms exhibit diverse details due to differences in sensor characteristics and imaging p

agentsarxiv-cs-cv
18 Aug 2026
Applications

SoftModel: A Neural Model That Grows Its Own Topology -- Governed Structural Growth for Continual In-Service Learning

DGX agent

arXiv:2608.16409v1 Announce Type: new Abstract: Today, a neural system is almost always used in two phases -- trained, then deployed -- and in that regime it freezes twice: training ends, and the topo

applicationsarxiv-cs-lg
18 Aug 2026
Research

TokenSTFormer: A Tokenized Spatial-temporal Attention Model for Holistic Motion Analysis in Adolescent Idiopathic Scoliosis Screening

DGX agent

arXiv:2608.16122v1 Announce Type: cross Abstract: Adolescent Idiopathic Scoliosis (AIS) is a prevalent spinal deformity in adolescents that, if left untreated, can result in severe health outcomes. Tr

researcharxiv-cs-ai
18 Aug 2026
Research

UC-VLM: Consistency-Driven Learning for AI-Generated Image Detection with Vision-Language Large Models

DGX agent

arXiv:2608.15238v1 Announce Type: new Abstract: Vision-Language Large Models (VLLMs) are promising for AI-generated image (AIGI) detection because they can produce both a prediction and a natural-lang

researcharxiv-cs-cv
18 Aug 2026
Tutorials

Variational Outlier-Robust Gaussian Process Regression with Generative Modeling

DGX agent

arXiv:2608.16606v1 Announce Type: new Abstract: Outliers can substantially distort Gaussian process regression (GPR) due to its conventional Gaussian observation likelihood, leading to inaccurate mode

tutorialsarxiv-cs-lg
18 Aug 2026
← Previous
1…254255256257258…1292
Next →