AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,403
  • Agents7,557
  • Applications5,411
  • Concepts5
  • Hardware1,836
  • Industry6,170
  • Local Ai4,931
  • Model Releases23,900
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,403
  • Agents7,557
  • Applications5,411
  • Concepts5
  • Hardware1,836
  • Industry6,170
  • Local Ai4,931
  • Model Releases23,900
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
88,403Total entries
1Added by human
88,402Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,623 results
Model Releases

Can Aha Moments Be Fake? Identifying True and Decorative Thinking Steps in Chain-of-Thought

DGX agent

arXiv:2510.24941v3 Announce Type: replace Abstract: Large language models can generate long chain-of-thought (CoT) reasoning, but it remains unclear whether the verbalized steps reflect the models' in

model-releasesarxiv-cs-lg
28 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Citation-Driven Multi-View Training for Patent Embeddings: QaECTER and Sophia-Bench

DGX agent

arXiv:2604.22897v1 Announce Type: cross Abstract: Patent retrieval underpins critical decisions in innovation, examination, and IP strategy, yet progress has been hampered by the absence of benchmarks

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Comparative Study of Weighted and Coupled Second- and Fourth-Order PDEs for Image Despeckling in Grayscale, Color, SAR, and Ultrasound

DGX agent

arXiv:2604.23612v1 Announce Type: new Abstract: Partial Differential Equation (PDE)-based approaches have gained significant attention in image despeckling due to their strong capability to preserve s

model-releasesarxiv-cs-cv
28 Apr 2026
Safety

DVPO: Distributional Value Modeling-based Policy Optimization for LLM Post-Training

DGX agent

arXiv:2512.03847v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has shown strong performance in LLM post-training, but real-world deployment often involves noisy or incomplete su

safetyarxiv-cs-ai
28 Apr 2026
Agents

Enabling Transparent Cyber Threat Intelligence Combining Large Language Models and Domain Ontologies

DGX agent

arXiv:2509.00081v2 Announce Type: replace-cross Abstract: Effective Cyber Threat Intelligence (CTI) relies upon accurately structured and semantically enriched information extracted from cybersecurity

agentsarxiv-cs-ai
28 Apr 2026
Research

Generalizable Friction Coefficient Estimation via Material Embedding and Proxy Interaction Modeling

DGX agent

arXiv:2604.24188v1 Announce Type: new Abstract: Accurately estimating friction coefficients between arbitrary material pairs is critical for robotics, digital fabrication, and physics-based simulation

researcharxiv-cs-ro
28 Apr 2026
Applications

Gradient-Guided Exploration of Generative Model's Latent Space for Controlled Iris Image Augmentations

DGX agent

arXiv:2511.09749v2 Announce Type: replace Abstract: Developing reliable iris recognition and presentation attack detection methods requires diverse datasets that capture realistic variations in iris f

applicationsarxiv-cs-cv
28 Apr 2026
Model Releases

JudgeSense: A Benchmark for Prompt Sensitivity in LLM-as-a-Judge Systems

DGX agent

arXiv:2604.23478v1 Announce Type: new Abstract: Large language models are increasingly deployed as automated judges for evaluating other models, yet the stability of their verdicts under semantically

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Judging the Judges: A Systematic Evaluation of Bias Mitigation Strategies in LLM-as-a-Judge Pipelines

DGX agent

arXiv:2604.23178v1 Announce Type: new Abstract: LLM-as-a-Judge has become the dominant paradigm for evaluating language model outputs, yet LLM judges exhibit systematic biases that compromise evaluati

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

MEASER: Malware embedding attacks on open-source LLMs

DGX agent

arXiv:2510.10486v2 Announce Type: replace-cross Abstract: Open-source large language models (LLMs) have demonstrated considerable dominance over proprietary LLMs in resolving neural processing tasks,

model-releasesarxiv-cs-ai
28 Apr 2026
Local Ai

Resource-Constrained UAV-Based Weed Detection for Site-Specific Management on Edge Devices

DGX agent

arXiv:2604.23442v1 Announce Type: new Abstract: Weeds compete with crops for light, water, and nutrients, reducing yield and crop quality. Efficient weed detection is essential for site-specific weed

local-aiarxiv-cs-cv
28 Apr 2026
Model Releases

Scheming Ability in LLM-to-LLM Strategic Interactions

DGX agent

arXiv:2510.12826v2 Announce Type: replace-cross Abstract: As large language model (LLM) agents are deployed autonomously in diverse contexts, evaluating their capacity for strategic deception becomes

model-releasesarxiv-cs-ai
28 Apr 2026
Research

SPEAR-1: Scaling Beyond Robot Demonstrations via 3D Understanding

DGX agent

arXiv:2511.17411v2 Announce Type: replace-cross Abstract: Robotic Foundation Models (RFMs) hold great promise as generalist, end-to-end systems for robot control. Yet their ability to generalize acros

researcharxiv-cs-lg
28 Apr 2026
Research

Adversarial Co-Evolution of Malware and Detection Models: A Bilevel Optimization Perspective

DGX agent

arXiv:2604.22569v1 Announce Type: cross Abstract: Machine learning-based malware detectors are increasingly vulnerable to adversarial examples. Traditional defenses, such as one-shot adversarial train

researcharxiv-cs-lg
27 Apr 2026
Model Releases

False Feasibility in Variable Impedance MPC for Legged Locomotion

DGX agent

arXiv:2604.22251v1 Announce Type: new Abstract: Variable impedance model predictive control (MPC) formulations that treat joint stiffness as an instantaneous decision variable operate on a feasible se

model-releasesarxiv-cs-ro
27 Apr 2026
Model Releases

How LLMs Detect and Correct Their Own Errors: The Role of Internal Confidence Signals

DGX agent

arXiv:2604.22271v1 Announce Type: new Abstract: Large language models can detect their own errors and sometimes correct them without external feedback, but the underlying mechanisms remain unknown. We

model-releasesarxiv-cs-lg
27 Apr 2026
Model Releases

How Popsa used Amazon Nova to inspire customers with personalised title suggestions

DGX agent

In this post, we share how we applied Amazon Bedrock and the Amazon Nova family of models to reimagine our Title Suggestion feature. By combining metadata, computer vision, and retrieval-augmented gen

model-releasesaws-ml-blog
27 Apr 2026
Model Releases

microsoft/VibeVoice

DGX agent

microsoft/VibeVoice VibeVoice is Microsoft's Whisper-style audio model for speech-to-text, MIT licensed and with speaker diarization built into the model. Microsoft released it on January 21st, 2026 b

model-releasessimon-willison
27 Apr 2026
Research

RouteLMT: Learned Sample Routing for Hybrid LLM Translation Deployment

DGX agent

arXiv:2604.22520v1 Announce Type: new Abstract: Large Language Models (LLMs) have achieved remarkable performance in Machine Translation (MT), but deploying them at scale remains prohibitively expensi

researcharxiv-cs-cl
27 Apr 2026
Research

Verbal Confidence Saturation in 3-9B Open-Weight Instruction-Tuned LLMs: A Pre-Registered Psychometric Validity Screen

DGX agent

arXiv:2604.22215v1 Announce Type: cross Abstract: Verbal confidence elicitation is widely used to extract uncertainty estimates from LLMs. We tested whether seven instruction-tuned open-weight models

researcharxiv-cs-ai
27 Apr 2026
Model Releases

Why there is no cloud version for Qwen 3.6 27/35B?

DGX agent

The Qwen 3.6-27B and 35B models are designed as open-weight models that developers can run locally on their own hardware without requiring cloud services. Alibaba released a separate cloud-only produc

model-releasesr-ollama
25 Apr 2026
Research

Attention-based multiple instance learning for predominant growth pattern prediction in lung adenocarcinoma wsi using foundation models

DGX agent

arXiv:2604.21530v1 Announce Type: cross Abstract: Lung adenocarcinoma (LUAD) grading depends on accurately identifying growth patterns, which are indicators of prognosis and can influence treatment de

researcharxiv-cs-ai
24 Apr 2026
Tutorials

Basic syntax from speech: Spontaneous concatenation in unsupervised deep neural networks

DGX agent

arXiv:2305.01626v4 Announce Type: replace-cross Abstract: Computational models of syntax are predominantly text-based. Here we propose that the most basic first step in the evolution of syntax can be

tutorialsarxiv-cs-ai
24 Apr 2026
Tutorials

Design, Modelling and Experimental Evaluation of a Tendon-driven Wrist Abduction-Adduction Mechanism for an upper limb exoskeleton

DGX agent

arXiv:2604.20893v1 Announce Type: new Abstract: Wrist exoskeletons play a vital role in rehabilitation and assistive applications, yet conventional actuation mechanisms such as electric motors or pneu

tutorialsarxiv-cs-ro
24 Apr 2026
Model Releases

Intent Laundering: AI Safety Datasets Are Not What They Seem

DGX agent

arXiv:2602.16729v3 Announce Type: replace-cross Abstract: We systematically evaluate the quality of widely used adversarial safety datasets from two perspectives: in isolation and in practice. In isol

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

LiveVLM: Efficient Online Video Understanding via Streaming-Oriented KV Cache and Retrieval

DGX agent

arXiv:2505.15269v2 Announce Type: replace Abstract: Recent developments in Video Large Language Models (Video LLMs) have enabled models to process hour-long videos and exhibit exceptional performance.

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

llm 0.31

DGX agent

Release: llm 0.31 New GPT-5.5 OpenAI model: llm -m gpt-5.5. #1418 New option to set the text verbosity level for GPT-5+ OpenAI models: -o verbosity low. Values are low, medium, high. New option for se

model-releasessimon-willison
24 Apr 2026
Applications

Reasoning Primitives in Hybrid and Non-Hybrid LLMs

DGX agent

arXiv:2604.21454v1 Announce Type: cross Abstract: Reasoning in large language models is often treated as a monolithic capability, but its observed gains may arise from more basic operations. We study

applicationsarxiv-cs-ai
24 Apr 2026
Model Releases

SurgViVQA: Temporally-Grounded Video Question Answering for Surgical Scene Understanding

DGX agent

arXiv:2511.03325v3 Announce Type: replace Abstract: Video Question Answering (VideoQA) in the surgical domain aims to enhance intraoperative understanding by enabling AI models to reason over temporal

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

VG-CoT: Towards Trustworthy Visual Reasoning via Grounded Chain-of-Thought

DGX agent

arXiv:2604.21396v1 Announce Type: cross Abstract: The advancement of Large Vision-Language Models (LVLMs) requires precise local region-based reasoning that faithfully grounds the model's logic in act

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Welcome DeepSeek V4 Pro Max https://huggingface.co/deepseek-ai/DeepSeek-V4-Pro

DGX agent

DeepSeek V4 Pro Max is a large language model released by DeepSeek AI and made available on Hugging Face, representing an advancement in their model lineup. The announcement was made by Clem Delangue,

model-releasesclem-delangue--x
24 Apr 2026
Research

3D Smoke Scene Reconstruction Guided by Vision Priors from Multimodal Large Language Models

DGX agent

arXiv:2604.05687v2 Announce Type: replace Abstract: Reconstructing 3D scenes from smoke-degraded multi-view images is particularly difficult because smoke introduces strong scattering effects, view-de

researcharxiv-cs-cv
23 Apr 2026
Research

A Kinematic Framework for Evaluating Pinch Configurations in Robotic Hand Design without Object or Contact Models

DGX agent

arXiv:2604.20692v1 Announce Type: new Abstract: Evaluating the pinch capability of a robotic hand is important for understanding its functional dexterity. However, many existing grasp evaluation metho

researcharxiv-cs-ro
23 Apr 2026
Safety

Breaking the Assistant Mold: Modeling Behavioral Variation in LLM Based Procedural Character Generation

DGX agent

arXiv:2601.03396v2 Announce Type: replace Abstract: Procedural content generation has enabled vast virtual worlds through levels, maps, and quests, but large-scale character generation remains underex

safetyarxiv-cs-cl
23 Apr 2026
Applications

Cortex 2.0: Grounding World Models in Real-World Industrial Deployment

DGX agent

arXiv:2604.20246v1 Announce Type: cross Abstract: Industrial robotic manipulation demands reliable long-horizon execution across embodiments, tasks, and changing object distributions. While Vision-Lan

applicationsarxiv-cs-ai
23 Apr 2026
Model Releases

Early-Stage Product Line Validation Using LLMs: A Study on Semi-Formal Blueprint Analysis

DGX agent

arXiv:2604.20523v1 Announce Type: cross Abstract: We study whether Large Language Models (LLMs) can perform feature model analysis operations (AOs) directly on semi-formal textual blueprints, i.e., co

model-releasesarxiv-cs-ai
23 Apr 2026
Research

Fourier Weak SINDy: Spectral Test Function Selection for Robust Model Identification

DGX agent

arXiv:2604.20141v1 Announce Type: new Abstract: We introduce Fourier Weak SINDy, a minimal noise-robust and interpretable derivative-free equation learning method that combines weak-form sparse equati

researcharxiv-cs-lg
23 Apr 2026
Model Releases

MambaLiteUNet: Cross-Gated Adaptive Feature Fusion for Robust Skin Lesion Segmentation

DGX agent

arXiv:2604.20286v1 Announce Type: cross Abstract: Recent segmentation models have demonstrated promising efficiency by aggressively reducing parameter counts and computational complexity. However, the

model-releasesarxiv-cs-ai
23 Apr 2026
Agents

Onboard Wind Estimation for Small UAVs Equipped with Low-Cost Sensors: An Aerodynamic Model-Integrated Filtering Approach

DGX agent

arXiv:2604.20290v1 Announce Type: new Abstract: To enable autonomous wind estimation for energy-efficient flight in small unmanned aerial vehicles (UAVs), this study proposes a method that estimates f

agentsarxiv-cs-ro
23 Apr 2026
Model Releases

Pushed: DFlash implementation for llama-cpp. buun-llama-cpp/llama-server -m Qwen3.6-27B.gguf -md dflash-draft-q4_k_m.gguf --spec-type dflash

DGX agent

This post demonstrates a DFlash implementation integrated with llama-cpp, showcasing a speculative decoding setup that uses Qwen 3.6-27B as the main model with a smaller draft model (dflash-draft-q4_k

model-releasesclem-delangue--x
23 Apr 2026
Model Releases

Scaling Self-Play with Self-Guidance

DGX agent

arXiv:2604.20209v1 Announce Type: new Abstract: LLM self-play algorithms are notable in that, in principle, nothing bounds their learning: a Conjecturer model creates problems for a Solver, and both i

model-releasesarxiv-cs-lg
23 Apr 2026
Model Releases

SGAP-Gaze: Scene Grid Attention Based Point-of-Gaze Estimation Network for Driver Gaze

DGX agent

arXiv:2604.19888v1 Announce Type: new Abstract: Driver gaze estimation is essential for understanding the driver's situational awareness of surrounding traffic. Existing gaze estimation models use dri

model-releasesarxiv-cs-cv
23 Apr 2026
Model Releases

Supplement Generation Training for Enhancing Agentic Task Performance

DGX agent

arXiv:2604.20727v1 Announce Type: cross Abstract: Training large foundation models for agentic tasks is increasingly impractical due to the high computational costs, long iteration cycles, and rapid o

model-releasesarxiv-cs-ai
23 Apr 2026
Agents

A Gesture-Based Visual Learning Model for Acoustophoretic Interactions using a Swarm of AcoustoBots

DGX agent

arXiv:2604.19643v1 Announce Type: new Abstract: AcoustoBots are mobile acoustophoretic robots capable of delivering mid-air haptics, directional audio, and acoustic levitation, but existing implementa

agentsarxiv-cs-ro
22 Apr 2026
Model Releases

Cyber Defense Benchmark: Agentic Threat Hunting Evaluation for LLMs in SecOps

DGX agent

arXiv:2604.19533v1 Announce Type: cross Abstract: We introduce the Cyber Defense Benchmark, a benchmark for measuring how well large language model (LLM) agents perform the core SOC analyst task of th

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

GeoLaux: A Benchmark for Evaluating MLLMs' Geometry Performance on Long-Step Problems Requiring Auxiliary Lines

DGX agent

arXiv:2508.06226v2 Announce Type: replace Abstract: Geometry problem solving (GPS) poses significant challenges for Multimodal Large Language Models (MLLMs) in diagram comprehension, knowledge applica

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

llama-server -hf ggml-org/Qwen3.6-27B-GGUF --spec-default

DGX agent

This post likely demonstrates running Qwen2 3.6B or 27B model in GGUF format using llama-server with default specifications, showcasing inference capabilities of quantized open-source models. The comm

model-releasesclem-delangue--x
22 Apr 2026
Model Releases

Multi-modal Test-time Adaptation via Adaptive Probabilistic Gaussian Calibration

DGX agent

arXiv:2604.19093v1 Announce Type: cross Abstract: Multi-modal test-time adaptation (TTA) enhances the resilience of benchmark multi-modal models against distribution shifts by leveraging the unlabeled

model-releasesarxiv-cs-ai
22 Apr 2026
← Previous
1…313314315316317…1326
Next →