AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

Content type
87,042Total entries
1Added by human
87,041Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,106 results
Model Releases

Robust and Personalized Federated Learning for Aircraft-Engine Prognostics under Benign and Adversarial Client Heterogeneity

DGX agent

arXiv:2608.04045v1 Announce Type: cross Abstract: Federated learning (FL) enables aircraft fleet operators to jointly train remaining-useful-life (RUL) models from engine sensor telemetry without shar

model-releasesarxiv-cs-ai
6 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Semantic Frame Interpolation

DGX agent

arXiv:2507.05173v2 Announce Type: replace Abstract: Generating intermediate video content of varying lengths based on given first and last frames, along with text prompt information, offers significan

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

Strengthening Target-Language Features: SAE-Based Steering for Multilingual Inference

DGX agent

arXiv:2608.04904v1 Announce Type: new Abstract: Multilingual large language models exhibit substantial performance differences across languages, while existing adaptation methods often require paramet

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

AS-FedBridge: Pseudo-Spike Bridge Distillation for Heterogeneous ANN-SNN Federated Learning

DGX agent

arXiv:2608.03324v1 Announce Type: new Abstract: Federated learning enables collaborative model training across distributed edge devices while strictly preserving data privacy. To facilitate practical

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

Balancing Efficiency and Efficacy: Training-Free Attention-Guided Switching Between Explicit and Latent Thoughts for MLLMs

DGX agent

arXiv:2608.03450v1 Announce Type: cross Abstract: Reasoning in Multimodal Large Language Models (MLLMs) requires both fine-grained visual perception and rigorous logical deduction. Explicit text-based

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Beyond Initialization Loss: A Systematic Study of Token Embedding Initialization Strategies for LLM Vocabulary Extension

DGX agent

arXiv:2608.03494v1 Announce Type: new Abstract: Vocabulary extension is an efficient way to adapt pretrained large language models (LLMs) to new languages, but the initialization of newly added token

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

Beyond Representational Similarity: Source-Conditioned Description-Length Gain for Generative Plagiarism Detection and Candidate Source Reranking

DGX agent

arXiv:2608.03859v1 Announce Type: cross Abstract: Large language models (LLMs) pose challenges to academic integrity and peer review. Yet generative plagiarism detection remains an underexplored and l

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Detecting Hallucinations and Recovering Verified Answers in Arabic Islamic Question Answering

DGX agent

arXiv:2608.03720v1 Announce Type: new Abstract: Large language models can generate fluent responses to Islamic questions while introducing factual errors that are difficult to identify. This paper pre

model-releasesarxiv-cs-cl
5 Aug 2026
Research

Earth Embeddings

DGX agent

arXiv:2608.03410v1 Announce Type: new Abstract: Earth observation is moving from foundation models that users must run themselves toward embedding products that package model feature outputs as reusab

researcharxiv-cs-cv
5 Aug 2026
Model Releases

Efficient Knowledge Distillation for LLMs: Offline Top-K Logits and a Fused Chunked KL Loss

DGX agent

arXiv:2608.03796v1 Announce Type: cross Abstract: Small language models are often the only option for deployment under tight latency, cost, and on-premises constraints, but they are rarely trained fro

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Evaluating Counterfactual Sensitivity to Patient Information in Medication-Safety Reasoning

DGX agent

arXiv:2608.03028v1 Announce Type: new Abstract: Applying a valid medication-safety rule when its patient-specific conditions are not met can produce an incorrect decision. Existing medical evaluations

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

How Closely Do LLM Reviews Align with Human Peer Review?

DGX agent

arXiv:2608.03659v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to generate scientific reviews, yet existing evaluations rarely examine whether different providers

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Latent Reward Registers for Diffusion Preference Alignment

DGX agent

arXiv:2608.03929v1 Announce Type: cross Abstract: Aligning diffusion models with human preferences usually relies on a sparse terminal reward evaluated on the final generated samples, presenting a sev

model-releasesarxiv-cs-cv
5 Aug 2026
Safety

MeSS: City Mesh-Guided Outdoor Scene Generation with Cross-View Consistent Diffusion

DGX agent

arXiv:2508.15169v4 Announce Type: replace Abstract: Mesh models have become increasingly accessible for numerous cities; however, the lack of realistic textures restricts their application in virtual

safetyarxiv-cs-cv
5 Aug 2026
Model Releases

Omega-S: A Functional Resilience Index for LLM Fine-Tuning

DGX agent

arXiv:2608.03887v1 Announce Type: new Abstract: Fine-tuning a large language model on new data degrades what it previously learned. We present Omega-S, a drop-in penalty computed from the weight matri

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

OncoTriad-QA: A Patient-Level Radiology-Pathology-Genomics Benchmark for Pan-Cancer Reasoning

DGX agent

arXiv:2608.02615v1 Announce Type: cross Abstract: Cancer diagnosis and characterization require integrating complementary evidence from radiology, pathology, genomics, and clinical metadata. However,

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

PI-Mem: Pushing Long-Context Reasoning to 3.6M Tokens with Parallel-Iterative Memory

DGX agent

arXiv:2608.03048v1 Announce Type: cross Abstract: Long-context reasoning remains a critical bottleneck for large language models, as recent recurrent-memory approaches face two inherent challenges: se

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

SAKI: Score-Aware Low-Rank Key Indexing for Long-Context KV Retrieval

DGX agent

arXiv:2608.03228v1 Announce Type: new Abstract: Existing low rank KV cache methods preserve either model weights or key variance, neither of which directly reflects the attention scores used during in

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

SeaSlides: Semantic Abstraction Layer for Agentic Slide Generation

DGX agent

arXiv:2608.03298v1 Announce Type: new Abstract: Agentic presentation generation must preserve source content, maintain coherent visual design, render specialized objects, and produce usable artifacts.

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Shorter Reasoning, Earlier Answers? An Evaluation of Reasoning Interfaces

DGX agent

arXiv:2608.03401v1 Announce Type: cross Abstract: Large language models often reason at length before answering, increasing cost and latency. Prompts and trained settings can shorten this reasoning, b

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

TACT: Taxonomy-Aligned Post-Training for Pedagogically Adaptive English Tutoring

DGX agent

arXiv:2608.03952v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used to provide conversational practice for English-as-a-second-language (ESL) learners. Effective ESL tut

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

UniWorld-Design: From Pixel Generation to Layer-Native Design

DGX agent

arXiv:2608.03971v1 Announce Type: new Abstract: We introduce UniWorld-Design, a framework that redefines image generation from flat pixel synthesis to structured visual composition, with semantic RGBA

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

VIVID: A Culturally Grounded Benchmark Exposing the Figurative Language Gap in Vietnamese NLP

DGX agent

arXiv:2608.03095v1 Announce Type: new Abstract: We present VIVID (Vietnamese Idioms for Validation and Interpretation Depth), the first systematic benchmark for evaluating culturally grounded figurati

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

When Many Answers Are Valid, Voting Fails: Symbolic Verification for Best-of-K Causal Reasoning in LLMs

DGX agent

arXiv:2608.03506v1 Announce Type: new Abstract: Self-consistency assumes the most frequent answer among sampled reasoning traces is the most reliable, but this can fail in causal reasoning: samples of

model-releasesarxiv-cs-ai
5 Aug 2026
Tutorials

Aggregate-then-Calibrate for Human-centered Assessment with Theoretical Guarantees

DGX agent

arXiv:2608.02455v1 Announce Type: new Abstract: Human-centered assessment tasks, which are essential for systematic decision-making, rely heavily on human judgment and typically lack verifiable ground

tutorialsarxiv-cs-lg
4 Aug 2026
Model Releases

Auditable Release Control for Pedagogical Leakage in LLM Tutors

DGX agent

arXiv:2608.00515v1 Announce Type: cross Abstract: Large language model tutors can be correct and helpful yet disclose an answer or decisive reasoning before that disclosure is authorized. We formalize

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Beyond Accuracy: Auditing Spatial Provenance in Visual Token Pruning for OCR-Critical MLLM Inference

DGX agent

arXiv:2608.00077v1 Announce Type: new Abstract: Visual-token pruning is usually judged by answer quality at a fixed retention budget. For text-rich multimodal large language models (MLLMs), this proto

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Bridging the English-Arabic Medical Knowledge Gap: Targeted Low-Rank Adaptation via Causal Layer Selection

DGX agent

arXiv:2608.00207v1 Announce Type: new Abstract: Large Language Models (LLMs) perform strongly in English medical tasks but degrade substantially in Arabic, a gap widely attributed to limited training

model-releasesarxiv-cs-cl
4 Aug 2026
Research

Cloud-ScPO: Hidden-State Geometry for Semi-Supervised Preference Optimization in LLM Reasoning

DGX agent

arXiv:2608.01014v1 Announce Type: new Abstract: Preference optimization improves mathematical reasoning in large language models (LLMs), but reliable chosen-rejected pairs usually require verified ans

researcharxiv-cs-cl
4 Aug 2026
Model Releases

CRIP: Channel Level Representation Injection for Personalized One-Shot Federated Learning

DGX agent

arXiv:2608.02222v1 Announce Type: new Abstract: One-shot federated learning (OSFL) has emerged as a promising collaborative model learning framework with only a single round of communication, offering

model-releasesarxiv-cs-lg
4 Aug 2026
Safety

Cross-Branch Conflict as a Shield: Safeguarding Facial Identities in Unified Multimodal Image Editing

DGX agent

arXiv:2607.16898v2 Announce Type: replace-cross Abstract: Unified multimodal models (UMMs) have recently demonstrated powerful instruction-based image editing capabilities, while also raising serious

safetyarxiv-cs-cl
4 Aug 2026
Model Releases

Deep Learning for Cyber Threat Detection and Mitigation in Healthcare-IoT

DGX agent

arXiv:2608.00118v1 Announce Type: cross Abstract: Cybersecurity is a fundamental requirement for protecting wearable devices used in healthcare Internet of Things (H-IoT) systems. Security failures in

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Do Static Embeddings Add Value to Hybrid Dutch Retrieval?

DGX agent

arXiv:2608.02112v1 Announce Type: new Abstract: Embedding benchmarks measure standalone model quality, but they do not establish whether a low-cost retriever contributes complementary ranking informat

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Fast and Accurate Quotation Attribution in Literary Texts

DGX agent

arXiv:2608.02359v1 Announce Type: new Abstract: Attributing quotations to their speakers in literary texts remains an open challenge. Standard methods, which independently predict a speaker mention fo

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Floor, Ceiling, and the Fusion Gap: How Much of Crowd Reading Attention Can Machines Predict?

DGX agent

arXiv:2608.01704v1 Announce Type: cross Abstract: A benchmark score means nothing without knowing what a trivial method achieves and what the best possible method could achieve. We construct both boun

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

GraphIR: Architecture-Level Search States for LLM-Guided Neural Architecture Evolution

DGX agent

arXiv:2608.01633v1 Announce Type: new Abstract: Large language models (LLMs) enable neural architecture search (NAS) directly over executable neural network programs. However, code-level flexibility d

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

HarnessCompass: Guiding Automatic Harness Evolution toward Generalizable and Effective Agent Harnesses

DGX agent

arXiv:2608.01918v1 Announce Type: cross Abstract: Harness design plays a critical role in agent performance by shaping how large language models (LLMs) perceive, reason over, and act within executable

model-releasesarxiv-cs-cl
4 Aug 2026
Tutorials

iMontage: Unified, Versatile, Highly Dynamic Many-to-many Image Generation

DGX agent

arXiv:2511.20635v3 Announce Type: replace Abstract: Pre-trained video models learn powerful priors for generating high-quality, temporally coherent content. While these models excel at temporal cohere

tutorialsarxiv-cs-cv
4 Aug 2026
Model Releases

Interpretable machine learning for predicting splitting strength of asphalt concrete: insights from SHAP analysis

DGX agent

arXiv:2608.00956v1 Announce Type: new Abstract: This paper presents an interpretable machine-learning framework for predicting the splitting strength (ST) of asphalt concrete and supporting data-drive

model-releasesarxiv-cs-lg
4 Aug 2026
Research

Length Penalties Make Chain-of-Thought Less Monitorable

DGX agent

arXiv:2607.09786v3 Announce Type: replace-cross Abstract: To curb overthinking and reduce inference costs, researchers now train reasoning models with penalties on chain of thought length. We find tha

researcharxiv-cs-cl
4 Aug 2026
Model Releases

Lethe: How Hard Is It to Forget? A Benchmark for Federated Unlearning in Medical Imaging

DGX agent

arXiv:2608.01094v1 Announce Type: new Abstract: Federated learning enables medical-imaging models to be trained across hospitals, and privacy law, most explicitly the GDPR ``right to be forgotten'', t

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

LongCat Sparse Attention: Taming the Lightning via Streaming-aware Hierarchical Cross-Layer Indexing

DGX agent

arXiv:2608.01662v1 Announce Type: cross Abstract: DeepSeek Sparse Attention (DSA) enables efficient long-context modeling through its Lightning Indexer. However, practical deployment remains constrain

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

LongChart VQA: A Comprehensive Benchmark for MLLMs with Complex Multi-Chart Reasoning

DGX agent

arXiv:2608.01328v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) are rapidly evolving with expanded context windows and stronger reasoning capabilities, enabling multi-chart un

model-releasesarxiv-cs-cl
4 Aug 2026
Safety

OSMDA: OpenStreetMap-based Domain Adaptation for Remote Sensing VLMs

DGX agent

arXiv:2603.11804v3 Announce Type: replace Abstract: Vision-Language Models (VLMs) adapted to remote sensing rely heavily on domain-specific image-text supervision, yet high-quality annotations for sat

safetyarxiv-cs-cv
4 Aug 2026
Model Releases

Protocol generalisation for brain tissue microstructure estimation via hypernetwork-controlled geometric deep learning

DGX agent

arXiv:2608.02053v1 Announce Type: cross Abstract: Brain tissue microstructure estimation with machine learning provides higher computational efficiency than conventional fitting. However, machine lear

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

Real-Time Visual Obstruction Detection in Surgical Augmented Reality

DGX agent

arXiv:2608.00232v1 Announce Type: new Abstract: Surgical augmented reality (AR) can provide contextual guidance by overlaying virtual annotations, tool cues, and procedural information onto the surgic

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

RSRA: Training-Free Probing of Representation Sensitivity for Efficient LoRA Rank Allocation

DGX agent

arXiv:2607.09757v2 Announce Type: replace Abstract: Parameter-efficient fine-tuning enables large language models to adapt to downstream tasks with substantially lower computational and storage cost,

model-releasesarxiv-cs-cv
4 Aug 2026
Local Ai

SpatialAfford: Teaching Compact VLMs Where to Look and Where to Ground for Affordance

DGX agent

arXiv:2608.00502v1 Announce Type: new Abstract: Affordance grounding aims to localize the functional region for interaction, such as the handle to grasp or the button to press, rather than the whole o

local-aiarxiv-cs-cv
4 Aug 2026
← Previous
1…340341342343344…1065
Next →