AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

Content type
AllBlog
87,042Total entries
1Added by human
87,041Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,509 results
Agents

DinoRADE: Full Spectral Radar-Camera Fusion with Vision Foundation Model Features for Multi-class Object Detection in Adverse Weather

DGX agent

arXiv:2604.08074v1 Announce Type: new Abstract: Reliable and weather-robust perception systems are essential for safe autonomous driving and typically employ multi-modal sensor configurations to achie

agentsarxiv-cs-cv
10 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Do MLLMs Really Understand Space? A Mathematical Reasoning Evaluation

DGX agent

arXiv:2602.11635v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have achieved strong performance on perception-oriented tasks, yet their ability to perform mathematical sp

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

ETCH-X: Robustify Expressive Body Fitting to Clothed Humans with Composable Datasets

DGX agent

arXiv:2604.08548v1 Announce Type: new Abstract: Human body fitting, which aligns parametric body models such as SMPL to raw 3D point clouds of clothed humans, serves as a crucial first step for downst

model-releasesarxiv-cs-cv
10 Apr 2026
Safety

Event-Centric World Modeling with Memory-Augmented Retrieval for Embodied Decision-Making

DGX agent

arXiv:2604.07392v1 Announce Type: cross Abstract: Autonomous agents operating in dynamic and safety-critical environments require decision-making frameworks that are both computationally efficient and

safetyarxiv-cs-ro
10 Apr 2026
Model Releases

EVGeoQA: Benchmarking LLMs on Dynamic, Multi-Objective Geo-Spatial Exploration

DGX agent

arXiv:2604.07070v1 Announce Type: new Abstract: While Large Language Models (LLMs) demonstrate remarkable reasoning capabilities, their potential for purpose-driven exploration in dynamic geo-spatial

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

FinTruthQA: A Benchmark for AI-Driven Financial Disclosure Quality Assessment in Investor -- Firm Interactions

DGX agent

arXiv:2406.12009v5 Announce Type: replace Abstract: Accurate and transparent financial information disclosure is essential for market efficiency, investor decision-making, and corporate governance. Ch

model-releasesarxiv-cs-cl
10 Apr 2026
Local Ai

Got early access to a real-time interactive video model, here's what I found

DGX agent

I was unable to retrieve the specific Reddit post at the provided URL through my search. The post (reddit.com/r/StableDiffusion/comments/1shxmfk) did not surface in the search results, and I cannot...

local-air-stablediffusion
10 Apr 2026
Agents

@hwchase17 middleware was the right abstraction for it too. way more adoptable than asking everyone to restructure their agent setup

DGX agent

LangChain's Middleware abstraction, introduced by Harrison Chase (@hwchase17) in LangChain 1.0 Alpha, addresses context engineering in AI agents by providing clean `before_model`, `after_model`, an...

agentsharrison-chase--x
10 Apr 2026
Safety

Improving Semantic Uncertainty Quantification in Language Model Question-Answering via Token-Level Temperature Scaling

DGX agent

arXiv:2604.07172v1 Announce Type: new Abstract: Calibration is central to reliable semantic uncertainty quantification, yet prior work has largely focused on discrimination, neglecting calibration. As

safetyarxiv-cs-lg
10 Apr 2026
Applications

Luwen Technical Report

DGX agent

arXiv:2604.06737v1 Announce Type: cross Abstract: Large language models have demonstrated remarkable capabilities across a wide range of natural language processing tasks, yet their application in the

applicationsarxiv-cs-ai
10 Apr 2026
Model Releases

ParseBench: A Document Parsing Benchmark for AI Agents

DGX agent

arXiv:2604.08538v1 Announce Type: new Abstract: AI agents are changing the requirements for document parsing. What matters is semantic correctness: parsed output must preserve the structure and

model-releasesarxiv-cs-cv
10 Apr 2026
Research

Physics-Informed Spectral Modeling for Hyperspectral Imaging

DGX agent

arXiv:2508.21618v2 Announce Type: replace-cross Abstract: We present PhISM, a physics-informed deep learning architecture that learns without supervision to explicitly disentangle hyperspectral observ

researcharxiv-cs-ai
10 Apr 2026
Local Ai

Q-Zoom: Query-Aware Adaptive Perception for Efficient Multimodal Large Language Models

DGX agent

arXiv:2604.06912v1 Announce Type: cross Abstract: MLLMs require high-resolution visual inputs for fine-grained tasks like document understanding and dense scene perception. However, current global res

local-aiarxiv-cs-ai
10 Apr 2026
Model Releases

Revisiting Radar Perception With Spectral Point Clouds

DGX agent

arXiv:2604.08282v1 Announce Type: new Abstract: Radar perception models are trained with different inputs, from range-Doppler spectra to sparse point clouds. Dense spectra are assumed to outperform sp

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

See: https://open.substack.com/pub/garymarcus/p/three-reasons-to-think-that-the-claude?r=8tdk6&utm_medium=ios

DGX agent

In an April 9, 2026 Substack post, AI skeptic Gary Marcus argues that Anthropic's announcement of its Claude 'Mythos' model was significantly overhyped, offering three key reasons for skepticism: t...

model-releasesgary-marcus--x
10 Apr 2026
Research

SPICE: Submodular Penalized Information-Conflict Selection for Efficient Large Language Model Training

DGX agent

arXiv:2601.23155v2 Announce Type: replace-cross Abstract: Information-based data selection for instruction tuning is compelling: maximizing the log-determinant of the Fisher information yields a monot

researcharxiv-cs-ai
10 Apr 2026
Safety

Synthetic Data for any Differentiable Target

DGX agent

arXiv:2604.08423v1 Announce Type: new Abstract: What are the limits of controlling language models via synthetic training data? We develop a reinforcement learning (RL) primitive, the Dataset Policy G

safetyarxiv-cs-cl
10 Apr 2026
Research

Through the Magnifying Glass: Adaptive Perception Magnification for Hallucination-Free VLM Decoding

DGX agent

arXiv:2503.10183v4 Announce Type: replace Abstract: Existing vision-language models (VLMs) often suffer from visual hallucination, where the generated responses contain inaccuracies that are not groun

researcharxiv-cs-cv
10 Apr 2026
Model Releases

TraceSafe: A Systematic Assessment of LLM Guardrails on Multi-Step Tool-Calling Trajectories

DGX agent

arXiv:2604.07223v1 Announce Type: cross Abstract: As large language models (LLMs) evolve from static chatbots into autonomous agents, the primary vulnerability surface shifts from final outputs to int

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Weakly Supervised Distillation of Hallucination Signals into Transformer Representations

DGX agent

arXiv:2604.06277v1 Announce Type: new Abstract: Existing hallucination detection methods for large language models (LLMs) rely on external verification at inference time, requiring gold answers, retri

model-releasesarxiv-cs-ai
10 Apr 2026
Research

Weaves, Wires, and Morphisms: Formalizing and Implementing the Algebra of Deep Learning

DGX agent

arXiv:2604.07242v1 Announce Type: new Abstract: Despite deep learning models running well-defined mathematical functions, we lack a formal mathematical framework for describing model architectures. Ad

researcharxiv-cs-lg
10 Apr 2026
Safety

What Drives Representation Steering? A Mechanistic Case Study on Steering Refusal

DGX agent

arXiv:2604.08524v1 Announce Type: cross Abstract: Applying steering vectors to large language models (LLMs) is an efficient and effective model alignment technique, but we lack an interpretable explan

safetyarxiv-cs-cl
10 Apr 2026
Model Releases

AI on the couch: Anthropic gives Claude 20 hours of psychiatry

DGX agent

As part of the evaluation of its Claude Mythos model, Anthropic engaged a clinical psychiatrist for approximately 20 hours of evaluation sessions, describing Mythos as 'the most psychologically se...

model-releasesars-technica
9 Apr 2026
Model Releases

With closed agent platforms like Claude Managed Agents, your agent's memory belongs to them, not you. It's locked behind their API. Agent in…

DGX agent

With closed agent platforms like Claude Managed Agents, your agent's memory belongs to them, not you. It's locked behind their API. Agent infra should be open: open harness, open memory, model agnosti

model-releasesharrison-chase--x
9 Apr 2026
Model Releases

Today, our group at @Mila_Quebec and the lab of @francesarnold at @Caltech just released a new paper I contributed to, exploring how multimo…

DGX agent

Today, our group at @Mila_Quebec and the lab of @francesarnold at @Caltech just released a new paper I contributed to, exploring how multimodal generative modeling could accelerate protein sciences! ⬇

model-releasesyoshua-bengio--x
8 Apr 2026
Industry

Frontier models will one shot just about anything few years Intelligence is compression Jevons law is more of a suggestion

DGX agent

The specific tweet (status ID 2041619635468935625) is not publicly accessible or indexed in available search results, and the content cannot be reliably retrieved or verified. I'm unable to produce...

industryemad-mostaque--x
7 Apr 2026
Model Releases

GLM 5.1 is now LIVE in Atomic Chat SOTA for code & chat – now runs locally with TurboQuant Thanks to @zai_org for open-sourcing this frontie…

DGX agent

GLM-5.1 is Z.ai's (zai-org) next-generation open-source flagship model for agentic engineering, achieving state-of-the-art performance on SWE-Bench Pro and significantly outperforming its predecess...

model-releaseszhipu-ai--x
7 Apr 2026
Hardware

Beyond FLOPs: Energy-Aware Knowledge Distillation for Sustainable LLMs on Code-Related Task

DGX agent

arXiv:2608.17515v1 Announce Type: cross Abstract: Background: Large Language Models (LLMs) are increasingly being applied to Software Engineering (SE) tasks, achieving high accuracy across problems su

hardwarearxiv-cs-ai
19 Aug 2026
Model Releases

Cross-Domain Generalization in Machine Unlearning via Label-Conditioned Energy Magnitude Regularization

DGX agent

arXiv:2608.17942v1 Announce Type: new Abstract: Machine unlearning removes the influence of specific data from a trained model. However, most methods treat the forgotten concept as isolated. In this p

model-releasesarxiv-cs-cv
19 Aug 2026
Model Releases

Dripper: Token-Efficient Main HTML Extraction with a Lightweight LM

DGX agent

arXiv:2511.23119v3 Announce Type: replace Abstract: High-quality main content extraction from web pages is a critical prerequisite for constructing large-scale training corpora. While traditional heur

model-releasesarxiv-cs-cl
19 Aug 2026
Model Releases

HarnessRisk: A Lifecycle-Oriented Benchmark for Agent Harness Safety

DGX agent

arXiv:2608.17597v1 Announce Type: cross Abstract: Large language models are increasingly deployed through agent harnesses that manage tools, extensions, persistent state, permissions, and external act

model-releasesarxiv-cs-ai
19 Aug 2026
Agents

It's felt like harness month on Twitter. We're seeing much faster and cheaper gains on a bunch of benchmarks by focusing on harness improvem…

DGX agent

It's felt like harness month on Twitter. We're seeing much faster and cheaper gains on a bunch of benchmarks by focusing on harness improvement rather than model improvement. Allowing the harness to b

agentsswyx--x
19 Aug 2026
Model Releases

@Kimi_Moonshot Kimi K3 is now available on Ollama's cloud subscriptions. We are working on improving Ollama's cloud to be much more transpar…

DGX agent

@Kimi_Moonshot Kimi K3 is now available on Ollama's cloud subscriptions. We are working on improving Ollama's cloud to be much more transparent on the pricing to show the best performance / $. Try it

model-releasesollama--x
19 Aug 2026
Model Releases

Replit Free Mode, powered by @OpenAI GPT-5.6 Luna, helps you maximize making while minimizing token costs.

DGX agent

Replit introduced **Free Mode** on August 19, 2026, powered by OpenAI’s GPT‑5.6 Luna model. The feature is designed to let users “maximise making while minimizing token costs,” offering free access wi

model-releasesreplit--x
19 Aug 2026
Model Releases

Thinking in a Low-Resource Language: What SFT Builds, What RL Fixes, What Accuracy Cannot See

DGX agent

arXiv:2608.17744v1 Announce Type: new Abstract: Take three frontier mixture-of-experts models (Alibaba, OpenAI, NVIDIA; 3.6-4.0B active parameters each) and fine-tune them to reason in a low-resource

model-releasesarxiv-cs-cl
19 Aug 2026
Model Releases

When Personalization Becomes Bias: Structural and Discursive Religious Framing in AI-Generated Financial Advice

DGX agent

arXiv:2608.16909v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly integrated into financial advisory systems, yet their role in reproducing religious bias remains underex

model-releasesarxiv-cs-ai
19 Aug 2026
Local Ai

When to Plan, When to Polish: Noise Level as a Granularity Axis for Diffusion Language Models

DGX agent

arXiv:2606.21802v2 Announce Type: replace Abstract: Standard tokenwise diffusion LMs keep training corruption and inference commitment at token granularity throughout denoising. At high noise, this le

local-aiarxiv-cs-cl
19 Aug 2026
Research

A cross-modal generative model for incomplete and degraded prostate MRI with multicentre clinical validation

DGX agent

arXiv:2608.16233v1 Announce Type: cross Abstract: Missing or degraded sequences can limit prostate multiparametric MRI. We developed MSCNet, a sequence-conditioned cross-modal generative framework for

researcharxiv-cs-ai
18 Aug 2026
Research

A Deep Learning Model for Spatially Clustered Data via Differentiable Cluster Assignment

DGX agent

arXiv:2608.14968v1 Announce Type: cross Abstract: We consider nonparametric regression when the association between a response and its covariates changes across an unknown partition of a spatial domai

researcharxiv-cs-lg
18 Aug 2026
Model Releases

AeroCopilotBench: A Two-Tier Benchmark for Evaluating LLM Agents as Aviation Copilots in an Interactive Virtual Cockpit Environment

DGX agent

arXiv:2608.16349v1 Announce Type: new Abstract: Large language model (LLM) agents may assist flight crews with complex decisions and task execution, but existing aviation evaluations centered on stati

model-releasesarxiv-cs-ai
18 Aug 2026
Model Releases

AeroGround: A Comprehensive Benchmark for Aerial-Ground Collaborative Reasoning

DGX agent

arXiv:2608.14721v1 Announce Type: new Abstract: Vision-language models (VLMs) have been widely employed in understanding and reasoning tasks for unmanned aerial vehicles (UAVs). Existing UAV benchmark

model-releasesarxiv-cs-cv
18 Aug 2026
Model Releases

Ask, Condition or Abstain: Reinforcement Learning for Missing-Premise Reasoning

DGX agent

arXiv:2608.16554v1 Announce Type: new Abstract: Answer-only reinforcement learning (RL) trains reasoning models to solve fully specified problems, but many realistic queries omit a premise needed for

model-releasesarxiv-cs-cl
18 Aug 2026
Model Releases

Auxiliary uncertainty signals for LLM-assisted systematic review screening: a benchmark across eight Cohen drug-class reviews

DGX agent

arXiv:2608.14551v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for title-abstract screening in systematic reviews, but their decisions lack calibrated uncertainty.

model-releasesarxiv-cs-cl
18 Aug 2026
Model Releases

Bye-bye, Bluebook? Automating Legal Drudgery With AI-Augmented Rule Following

DGX agent

arXiv:2505.02763v2 Announce Type: replace-cross Abstract: One of the central promises of legal AI is to automate drudgery -- the formal, repetitive tasks of lawyers' work that consume time without cal

model-releasesarxiv-cs-ai
18 Aug 2026
Model Releases

Can LLMs Reason Like Automated Theorem Provers for Rust Verification? VCoT-Bench: Evaluating via Verification Chain of Thought

DGX agent

arXiv:2603.18334v2 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) increasingly assist secure software development, their ability to meet the rigorous demands of Rust program ve

model-releasesarxiv-cs-ai
18 Aug 2026
Model Releases

Disentangling Pictorial Cue Understanding from Language Bias in VLMs via Depth Ordering Task

DGX agent

arXiv:2607.01503v2 Announce Type: replace Abstract: In this paper, we study depth perception of vision-language models (VLMs) to isolate the effects of pictorial depth cues and disentangle vision and

model-releasesarxiv-cs-cv
18 Aug 2026
Model Releases

Do LLM Agents Negotiate Rationally? A Mechanism-Design Framework for Verifiable Multi-Agent Interaction over A2A/MCP

DGX agent

arXiv:2608.14613v1 Announce Type: new Abstract: Modern LLM-agent frameworks increasingly interoperate through standards such as Anthropic's Model Context Protocol (MCP) for agent-to-tool access and Go

model-releasesarxiv-cs-ai
18 Aug 2026
Safety

Every Coin Has Two Sides: On the Dual Nature of Generalization in On-Policy Distillation of Large Language Models

DGX agent

arXiv:2608.16647v1 Announce Type: new Abstract: On-policy distillation (OPD) transfers teacher capabilities by supervising trajectories sampled from the student's own policy, yet its generalization be

safetyarxiv-cs-cl
18 Aug 2026
← Previous
1…338339340341342…1303
Next →