AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
85,202Total entries
1Added by human
85,201Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,811 results
Model Releases

Restoring Heterogeneity in LLM-based Social Simulation: An Audience Segmentation Approach

DGX agent

arXiv:2604.06663v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used to simulate social attitudes and behaviors, offering scalable 'silicon samples' that can approximat

model-releasesarxiv-cs-ai
10 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Severity-Aware Weighted Loss for Arabic Medical Text Generation

DGX agent

arXiv:2604.06346v1 Announce Type: cross Abstract: Large language models have shown strong potential for Arabic medical text generation; however, traditional fine-tuning objectives treat all medical ca

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Tabular GANs for uneven distribution

DGX agent

arXiv:2010.00638v2 Announce Type: replace-cross Abstract: Generative models for tabular data have evolved rapidly beyond Generative Adversarial Networks (GANs). While GANs pioneered synthetic tabular

model-releasesarxiv-cs-cv
10 Apr 2026
Research

Testimole-Conversational: A 30-Billion-Word Italian Discussion Board Corpus (1996-2024) for Language Modeling and Sociolinguistic Research

DGX agent

arXiv:2602.14819v2 Announce Type: replace Abstract: We present 'Testimole-conversational' a massive collection of discussion boards messages in the Italian language. The large size of the corpus, more

researcharxiv-cs-cl
10 Apr 2026
Model Releases

The AI Skills Shift: Mapping Skill Obsolescence, Emergence, and Transition Pathways in the LLM Era

DGX agent

arXiv:2604.06906v1 Announce Type: cross Abstract: As Large Language Models reshape the global labor market, policymakers and workers need empirical data on which occupational skills may be most suscep

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

The Theory and Practice of Highly Scalable Gaussian Process Regression with Nearest Neighbours

DGX agent

arXiv:2604.07267v1 Announce Type: cross Abstract: Gaussian process (GP) regression is a widely used non-parametric modeling tool, but its cubic complexity in the training size limits its use on mass

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

ToxReason: A Benchmark for Mechanistic Chemical Toxicity Reasoning via Adverse Outcome Pathway

DGX agent

arXiv:2604.06264v1 Announce Type: cross Abstract: Recent advances in large language models (LLMs) have enabled molecular reasoning for property prediction. However, toxicity arises from complex biolog

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Video Parallel Scaling: Aggregating Diverse Frame Subsets for VideoLLMs

DGX agent

arXiv:2509.08016v2 Announce Type: replace Abstract: Video Large Language Models (VideoLLMs) face a critical bottleneck: increasing the number of input frames to capture fine-grained temporal detail le

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

WASD: Locating Critical Neurons as Sufficient Conditions for Explaining and Controlling LLM Behavior

DGX agent

arXiv:2603.18474v2 Announce Type: replace Abstract: Precise behavioral control of large language models (LLMs) is critical for complex applications. However, existing methods often incur high training

model-releasesarxiv-cs-cl
10 Apr 2026
Applications

A Year in LLM Serving: Workload Evolution, Caching and Load-Balancing

DGX agent

arXiv:2608.13573v1 Announce Type: new Abstract: Large Language Model (LLM) serving has become a critical cloud workload, and realistic traces are essential for motivating and benchmarking serving syst

applicationsarxiv-cs-ai
17 Aug 2026
Model Releases

AnchorBench: A Multi-Pathway Benchmark for the Anchoring Effect in LLMs

DGX agent

arXiv:2608.14320v1 Announce Type: new Abstract: The anchoring effect is a cognitive bias in which an initial reference value shifts a later judgment toward itself. This effect is well established in h

model-releasesarxiv-cs-ai
17 Aug 2026
Research

Automated Inference of Graph Transformation Rules

DGX agent

arXiv:2404.02692v3 Announce Type: replace-cross Abstract: The explosion of data available in life sciences is fueling an increasing demand for expressive models and computational methods. Graph transf

researcharxiv-cs-lg
17 Aug 2026
Model Releases

Bootstrapping Niche Multilingual Code Translation via Reinforcement Learning with Execution-Based Verifiable Supervision

DGX agent

arXiv:2608.13854v1 Announce Type: new Abstract: Code translation must preserve executable behavior across many programming languages, yet neural code translation has largely focused on a few popular l

model-releasesarxiv-cs-cl
17 Aug 2026
Model Releases

ChartProbe: A Diagnostic Study on Visual Reasoning through Perception, Grounding, and Simple Reasoning

DGX agent

arXiv:2608.13766v1 Announce Type: new Abstract: Vision-language models (VLMs) remain unreliable on chart questions that require reasoning over visual quantities, and this weakness is usually attribute

model-releasesarxiv-cs-cv
17 Aug 2026
Research

Concept Guidance: Precise, Training-Free Latent Control for Text-to-Image Generation

DGX agent

arXiv:2608.14172v1 Announce Type: cross Abstract: Text-to-image diffusion models have two major drawbacks that severely limit their practical utility: (1) standard models lack an intrinsic mechanism f

researcharxiv-cs-ai
17 Aug 2026
Local Ai

CoViLLM: An Adaptive Human-Robot Collaborative Assembly Framework Using Large Language Models

DGX agent

arXiv:2603.11461v3 Announce Type: replace Abstract: With increasing demand for mass customization, traditional manufacturing robots that rely on rule-based operations lack the flexibility to accommoda

local-aiarxiv-cs-ro
17 Aug 2026
Model Releases

Don't Claim Benchmark-Oriented Optimization Improves General Coding Capability -- Diverse Evaluation Is Required

DGX agent

arXiv:2608.13566v1 Announce Type: cross Abstract: Post-training papers, model cards, and blog posts often treat scores on a small set of coding benchmarks (e.g., SWE-bench and LiveCodeBench) as eviden

model-releasesarxiv-cs-ai
17 Aug 2026
Model Releases

Nanbeige4.2-3B on Apple Silicon: Fixing Deployment Bugs and Decreasing Looped Transformer Memory Overhead

DGX agent

arXiv:2608.13987v1 Announce Type: new Abstract: Nanbeige4.2-3B is a 3B-parameter agentic model built around a Looped Transformer (LT) that reuses one stack of layers for a second forward pass, adding

model-releasesarxiv-cs-ai
17 Aug 2026
Research

Offline Deep Q* Estimation with Diffusion Models

DGX agent

arXiv:2608.14401v1 Announce Type: cross Abstract: In offline RL, estimating the optimal action-value function Q^* can be formulated as solving the optimal Bellman equation based solely on offline obse

researcharxiv-cs-lg
17 Aug 2026
Research

Ontology-Grounded World Models for Failure Diagnosis and Closed-Loop Repair in Physical AI Systems

DGX agent

arXiv:2608.13901v1 Announce Type: new Abstract: EV-WM represents candidate quality with feature and event scores, but these scores do not explicitly record an unmet task predicate, a route label for a

researcharxiv-cs-ro
17 Aug 2026
Model Releases

P2Skill: Privacy Preserving Skill Distillation for Cloud-Local LLM Inference Systems

DGX agent

arXiv:2608.14094v1 Announce Type: cross Abstract: Cloud-local LLM inference systems have the potential to use the reasoning capability of large cloud models while protecting sensitive user data on per

model-releasesarxiv-cs-ai
17 Aug 2026
Model Releases

QuaSAR: Quantization Compensation via Stable Activation-Aware Rank Truncation

DGX agent

arXiv:2608.14149v1 Announce Type: new Abstract: Recent training-free post-training quantization methods restore model accuracy through closed-form residual compensation. To constrain additional model

model-releasesarxiv-cs-ai
17 Aug 2026
Applications

Reading Between The Lines: Modeling and Evaluating Behavioral Realism in Legal Simulation

DGX agent

arXiv:2608.13712v1 Announce Type: cross Abstract: Deposition training requires attorneys to manage dynamic witness behavior, yet legal-AI evaluations largely focus on factual accuracy, reasoning, or r

applicationsarxiv-cs-ai
17 Aug 2026
Model Releases

Revisiting Shape and Texture Reliance with Category-Separability-Calibrated Suppression

DGX agent

arXiv:2607.16298v2 Announce Type: replace Abstract: Feature-suppression evaluations infer model reliance on shape or texture from the accuracy loss caused by attenuating each type of information. Such

model-releasesarxiv-cs-cv
17 Aug 2026
Applications

When Denoising Hurts: Rethinking the Terminal Step of Diffusion Time Series Forecasters -- Extended Version

DGX agent

arXiv:2608.14067v1 Announce Type: new Abstract: Diffusion models offer a natural way to model uncertainty in time series forecasting, yet their iterative sampling process is often treated as a uniform

applicationsarxiv-cs-lg
17 Aug 2026
Model Releases

BavGround: A Benchmark for Regional Cultural Grounding and Dialect Competence in Bavarian

DGX agent

arXiv:2608.12894v1 Announce Type: new Abstract: Cultural evaluation of large language models (LLMs) often focuses on high-resource standard languages, leaving regional culture and dialect communities

model-releasesarxiv-cs-cl
14 Aug 2026
Model Releases

CangjieBench: Benchmarking LLMs on a Low-Resource General-Purpose Programming Language

DGX agent

arXiv:2603.14501v2 Announce Type: replace-cross Abstract: Large Language Models excel in high-resource programming languages but struggle with low-resource ones. Existing research related to low-resou

model-releasesarxiv-cs-ai
14 Aug 2026
Research

CityRiSE: Reasoning Urban Socio-Economic Status in Large Vision-Language Models via Reinforcement Learning

DGX agent

arXiv:2510.22282v2 Announce Type: replace-cross Abstract: Urban socio-economic sensing plays a vital role in advancing global sustainable development goals. With the advent of Large Vision-Language Mo

researcharxiv-cs-ai
14 Aug 2026
Safety

CW-BASS v2: Saturation-Aware Pseudo-Label Selection for Semi-Supervised Segmentation under Foundation-Model Teachers

DGX agent

arXiv:2608.12773v1 Announce Type: new Abstract: Semi-supervised semantic segmentation has long turned on one question, which pseudo-labels to trust, and a generation of selection rules, dynamic thresh

safetyarxiv-cs-cv
14 Aug 2026
Model Releases

DFM Mimir v1: An Open HRM Delivering Frontier Performance at 1B Parameters Using Only Permissible Post-Training Data

DGX agent

arXiv:2608.13517v1 Announce Type: cross Abstract: Current large language model development relies on massive, often non-permissible datasets, creating a high barrier for researchers committed to open-

model-releasesarxiv-cs-ai
14 Aug 2026
Research

DMDIntel: Interpreting Large Language Models via Dynamic Mode Decomposition

DGX agent

arXiv:2608.13048v1 Announce Type: new Abstract: In this work, we introduce DMDIntel which uses dynamic mode decomposition (DMD) to make the predictions made by LLMs in a classification task interpreta

researcharxiv-cs-ai
14 Aug 2026
Model Releases

EgoMonth: A Month-Level Egocentric Video Benchmark for Long-Term Spatiotemporal Memory

DGX agent

arXiv:2608.13113v1 Announce Type: cross Abstract: Recent advances in Multimodal Large Language Models (MLLMs) have led to substantial progress in video understanding, accompanied by a growing number o

model-releasesarxiv-cs-ai
14 Aug 2026
Safety

Evaluation Resolution Confounds Learning-Rule Comparisons in Model-Brain RSA of Early Visual Cortex

DGX agent

arXiv:2608.12408v1 Announce Type: cross Abstract: Representational similarity analysis (RSA) is increasingly used to ask which learning rules give convolutional networks brain-like representations. Be

safetyarxiv-cs-lg
14 Aug 2026
Model Releases

Falsehood and Impossibility Are Different Directions in an AI's Representation of Language

DGX agent

arXiv:2608.12852v1 Announce Type: cross Abstract: Language can describe states of affairs that are false and states of affairs that could not be the case at all. Whether an AI model internally disting

model-releasesarxiv-cs-ai
14 Aug 2026
Research

GS^{2}CI: Robust Gaussian Splatting For Snapshot Compressive Imaging via Large Vision Model Priors

DGX agent

arXiv:2608.13502v1 Announce Type: new Abstract: Snapshot Compressive Imaging (SCI) offers an efficient solution for high-speed video acquisition and, under exposure-time camera--scene relative motion,

researcharxiv-cs-cv
14 Aug 2026
Model Releases

Refusing Intent, Not Form: Wrapper-Based Intent-Group Supervision for LLM Safety

DGX agent

arXiv:2608.13304v1 Announce Type: new Abstract: Safety tuning can improve harmful refusal, but models may learn surface-form shortcuts: wrapped harmful prompts bypass safety, while similarly wrapped b

model-releasesarxiv-cs-cl
14 Aug 2026
Model Releases

Which LLM Is Your Ideal Companion? Evaluating Emotional Companion Capabilities of LLMs Based on Adult Attachment Theory

DGX agent

arXiv:2608.13168v1 Announce Type: new Abstract: As large language models (LLMs) are increasingly applied for emotional companionship, evaluating their behavior and capabilities in intimate relationshi

model-releasesarxiv-cs-cl
14 Aug 2026
Model Releases

Beyond Parameter Space: NTK-Guided Personalized Aggregation for Robust Federated Learning

DGX agent

arXiv:2608.12108v1 Announce Type: new Abstract: Federated learning (FL) enables collaborative model training across distributed clients while keeping data local. A central challenge is determining whi

model-releasesarxiv-cs-lg
13 Aug 2026
Model Releases

COGENT: Counterfactual Gaussian Explanations for Volumetric Medical Images

DGX agent

arXiv:2608.11422v1 Announce Type: new Abstract: Explainability is essential for deploying deep learning models in high-stakes medical applications. Existing explainability methods for volumetric imagi

model-releasesarxiv-cs-cv
13 Aug 2026
Research

Decoupled Quadratic Kalman Filter for Elliptical Extended Object Tracking with Log-normal Axis Modeling

DGX agent

arXiv:2512.14426v2 Announce Type: replace-cross Abstract: Extended object tracking involves estimating both the physical extent and kinematic parameters of a target object, where typically multiple me

researcharxiv-cs-ro
13 Aug 2026
Safety

LabelFusion-TS: Fusing Large Language Models, Transformer Encoders, and Financial Time Series for Monetary-Policy Stance Classification

DGX agent

arXiv:2608.11753v1 Announce Type: new Abstract: Financial text is produced and interpreted within a market environment, yet financial text classifiers almost always receive text alone. We study whethe

safetyarxiv-cs-cl
13 Aug 2026
Research

Massive Activations in Hybrid Linear Attention Large Language Models: Pre-Attention Spikes and Inter-Spike Plateaus

DGX agent

arXiv:2608.12149v1 Announce Type: new Abstract: We present the first systematic study of Massive activations (MAs) in layer-interleaved HLA LLMs and uncover two architecture-aligned morphologies: MAs

researcharxiv-cs-cl
13 Aug 2026
Research

ScreenShot: A Foundation Model for Few-Shot Combination Drug Screening

DGX agent

arXiv:2608.12219v1 Announce Type: new Abstract: Treating patients with combinations of drugs reduces the risk of resistance to any individual drug. Finding effective combinations is difficult because

researcharxiv-cs-lg
13 Aug 2026
Model Releases

Self-Evolving Code-with-Image Reasoning

DGX agent

arXiv:2608.11292v1 Announce Type: new Abstract: Multimodal models increasingly reach for tools when solving visual tasks (crop, zoom, rotate, brighten), a paradigm known as thinking-with-images. The c

model-releasesarxiv-cs-cv
13 Aug 2026
Model Releases

TangPoetryBench: A Multi-Dimensional Benchmark and Rubric-Conditioned Evaluator for Poetry-to-Image Generation

DGX agent

arXiv:2608.11452v1 Announce Type: cross Abstract: Text-to-image (T2I) models are increasingly asked to illustrate literary and cultural content, yet we cannot measure how well an image renders the mea

model-releasesarxiv-cs-ai
13 Aug 2026
Research

Towards foundation-style models for energy-frontier heterogeneous neutrino detectors via self-supervised pre-training

DGX agent

arXiv:2604.07037v2 Announce Type: replace-cross Abstract: Accelerator-based neutrino physics is entering an energy-frontier regime in which interactions reach the TeV scale and produce exceptionally d

researcharxiv-cs-cv
13 Aug 2026
Research

Towards Model-based Run-time Cybersecurity: On Control-Flow Anomaly Detection, Attack Identification, and Hardware Monitoring

DGX agent

arXiv:2608.11802v1 Announce Type: cross Abstract: Methods to increase the resilience of systems to cyber-attacks become increasingly important. Control-flow monitoring provides a principled basis to e

researcharxiv-cs-ai
13 Aug 2026
Model Releases

When the API Speaks the Wrong Language: Revisiting Post-Training for Multilingual Tool Use

DGX agent

arXiv:2608.11715v1 Announce Type: cross Abstract: The reliability of Large Language Models (LLMs) for API calling degrades in multilingual settings. A common failure occurs when a model selects the co

model-releasesarxiv-cs-ai
13 Aug 2026
← Previous
1…244245246247248…1038
Next →