AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlog
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,646 results
Model Releases

AutoSupervision: Closing the Feedback Loop in Scientific Workflows with Grounded Revision Verification

DGX agent

arXiv:2607.27845v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have enabled AI systems to assist scientific research and peer review. However, an essential capability

model-releasesarxiv-cs-cl
31 Jul 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

AWARE-FX: An Auditable Knowledge-Guided AI System for Measuring Corporate Foreign-Exchange Hedging Disclosure

DGX agent

arXiv:2607.27611v1 Announce Type: new Abstract: Corporate annual reports contain weakly structured evidence about foreign-exchange risk management, derivative use, natural hedging, and explicit non-us

researcharxiv-cs-cl
31 Jul 2026
Model Releases

Baikal: Structured Search for Deep Research over Data Lakes

DGX agent

arXiv:2607.27726v1 Announce Type: cross Abstract: Deep research over data lakes requires an LLM agent to investigate evidence across thousands of heterogeneous tables and passages to synthesize a repo

model-releasesarxiv-cs-cl
31 Jul 2026
Agents

Belief Coevolution in a Social Network of Generalist and Specialist Large Language Models

DGX agent

arXiv:2607.27512v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed in multi-agent environments. However, the processes by which beliefs form and propagate among int

agentsarxiv-cs-cl
31 Jul 2026
Model Releases

Benchmarking LLM Competence on Logical Inference over Probability Operators

DGX agent

arXiv:2607.27405v1 Announce Type: new Abstract: Both expressions of uncertainty and inferences are ubiquitous in natural language, and valid inferences over natural-language expressions of uncertainty

model-releasesarxiv-cs-cl
31 Jul 2026
Research

Beyond a Single Judge: Simulating Social Persona Panels for Generative UI Evaluation

DGX agent

arXiv:2607.28439v1 Announce Type: new Abstract: Generative UI (GenUI) lets large language models synthesize a complete, renderable interface directly from a natural-language instruction, but evaluatin

researcharxiv-cs-cl
31 Jul 2026
Model Releases

Beyond Borrowed Histories: Person-Aligned User Simulation for Interactive Role-Playing Evaluation

DGX agent

arXiv:2607.27816v1 Announce Type: new Abstract: Role-playing agents (RPAs) have become one of the most important consumer applications of large language models. Users engage in multi-turn conversation

model-releasesarxiv-cs-cl
31 Jul 2026
Safety

Beyond Feeling Better: Capability-Sustaining Emotional Dialogue as a Longitudinal Research Paradigm

DGX agent

arXiv:2607.27851v1 Announce Type: new Abstract: Emotional dialogue research includes two influential strategy traditions. Empathetic dialogue prioritizes understanding a speaker's emotional experience

safetyarxiv-cs-cl
31 Jul 2026
Model Releases

Beyond Sentiment: Structured Information Extraction from Financial News

DGX agent

arXiv:2607.28496v1 Announce Type: new Abstract: Financial sentiment analysis has become a standard component in news-driven stock prediction, yet it reduces rich, multi-dimensional news articles to a

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Beyond Similarity: Grounded Agentic Extraction and Expert-Adjudicated Evaluation of Intertextuality in Classical Chinese Histories

DGX agent

arXiv:2607.27595v1 Announce Type: new Abstract: Computational approaches to intertextuality have advanced from string matching to neural retrieval, yet their outputs, similarity scores and parallel-pa

model-releasesarxiv-cs-cl
31 Jul 2026
Safety

BridgeAlign: Bridging Preference Alignment for Humanities and Social Sciences

DGX agent

arXiv:2607.27366v1 Announce Type: new Abstract: While data synthesis for large language models (LLMs) is prevalent, it primarily targets domains with verifiable answers, overlooking open-ended humanit

safetyarxiv-cs-cl
31 Jul 2026
Applications

CACHE-UK: A Stability-Aware Memory Editor for Sequentially Updated Quantized LLMs in Finance

DGX agent

arXiv:2607.28292v1 Announce Type: new Abstract: Large Language Models (LLMs) deployed in dynamic financial environments face a critical challenge: maintaining factual accuracy as market conditions, re

applicationsarxiv-cs-cl
31 Jul 2026
Model Releases

Can Agents Deceive? Evaluating Reasoning and Deception in ParliamentBench using a Social Deduction Game

DGX agent

arXiv:2607.28146v1 Announce Type: new Abstract: As large language models (LLMs) are deployed as agents in high-stakes settings, such as medical and legal systems, understanding their deceptive capabil

model-releasesarxiv-cs-cl
31 Jul 2026
Tutorials

Can Large Language Models Execute Parent Orders?

DGX agent

arXiv:2607.28410v1 Announce Type: cross Abstract: Parent-order execution is a core problem in algorithmic trading, where the goal is to split a large order into smaller orders while reducing execution

tutorialsarxiv-cs-cl
31 Jul 2026
Model Releases

Can LVLMs Uncover the Truth Behind Visual Illusions? An Analysis of Perceptual and Reasoning Capabilities

DGX agent

arXiv:2607.27747v1 Announce Type: new Abstract: Large Vision Language Models have integrated reasoning capabilities, elevating cognitive performance to new levels. However, existing evaluations either

model-releasesarxiv-cs-cl
31 Jul 2026
Research

Causal Discovery with Inverted Self-attention for Multivariate Time Series

DGX agent

arXiv:2607.28212v1 Announce Type: new Abstract: Causal discovery in multivariate time series data is challenging due to complex interactions, high dimensionality, and nonlinear dependencies among vari

researcharxiv-cs-cl
31 Jul 2026
Tutorials

CDAE: Enhancing Perturbation Robustness in Pretrained Language Models with Contrastive Denoising

DGX agent

arXiv:2607.28236v1 Announce Type: cross Abstract: Pre-trained language models have significantly improved sentence representation learning, yet their embedding remain sensitive to semantic preserving

tutorialsarxiv-cs-cl
31 Jul 2026
Applications

Challenges in annotations by humans and LLMs: A case study of evaluative language

DGX agent

arXiv:2607.28119v1 Announce Type: new Abstract: In this paper, we draw a comparison between linguists in training, a trained linguist, and annotations generated by large language models (LLMs) to find

applicationsarxiv-cs-cl
31 Jul 2026
Agents

Change2Task: From Repository Changes to Executable Coding Agent Tasks and Environments

DGX agent

arXiv:2607.28591v1 Announce Type: cross Abstract: Scaling coding agents requires a continuing supply of executable data for training, benchmarking, and continuous evaluation. Each task must couple a r

agentsarxiv-cs-cl
31 Jul 2026
Model Releases

ChronoMem: Version Control and Semantic Rollback for Large Language Model Agent Memory

DGX agent

arXiv:2607.27773v1 Announce Type: new Abstract: LLM agents increasingly rely on long-term memory to support multi-session interaction and personalization. However, existing agent memory systems are de

model-releasesarxiv-cs-cl
31 Jul 2026
Applications

Cocktail-Talker: Multi-Speaker Dialog Modeling in Noisy Social Environments with Turn Action GRPO

DGX agent

arXiv:2607.27756v1 Announce Type: cross Abstract: Spoken dialog systems are typically designed for clean, dyadic interactions in which a single user and an assistant take turns speaking. Real-world so

applicationsarxiv-cs-cl
31 Jul 2026
Applications

Correlation between prosody and pragmatics: A case study of the discourse marker hala `now' in Persian

DGX agent

arXiv:2607.28359v1 Announce Type: new Abstract: The Persian discourse marker hala ('now') exhibits remarkable multifunctionality, extending far beyond its temporal adverbial role to encompass a variet

applicationsarxiv-cs-cl
31 Jul 2026
Safety

Creative Transformation in Literary Texts: Modelling Change Across Representational Levels

DGX agent

arXiv:2607.28513v1 Announce Type: new Abstract: Creativity is often framed as the production of novelty, yet many cultural works emerge through transformation of earlier artifacts and not through isol

safetyarxiv-cs-cl
31 Jul 2026
Agents

CRMWeaver: Building Powerful Business Agent via Agentic RL and Shared Memories

DGX agent

arXiv:2510.25333v2 Announce Type: replace Abstract: Recent years have witnessed the rapid development of LLM-based agents, which shed light on using language agents to solve complex real-world problem

agentsarxiv-cs-cl
31 Jul 2026
Safety

Digital Harf: A Clinically Integrated Multimodal AI System for Pervasive Arabic Speech and Language Therapy

DGX agent

arXiv:2607.27212v1 Announce Type: cross Abstract: Children with Autism Spectrum Disorder in Arabic-speaking countries face compounded barriers to effective speech and language therapy: a shortage of q

safetyarxiv-cs-cl
31 Jul 2026
Model Releases

Dimensionality and Measurement Precision in HLE's Multiple-Choice Subset

DGX agent

arXiv:2607.27420v1 Announce Type: cross Abstract: Humanity's Last Exam (HLE) is widely used to evaluate frontier language models. HLE organizes its questions into eight subject-domain categories, whos

model-releasesarxiv-cs-cl
31 Jul 2026
Safety

DualAnchor: Preserving Language Priors and Improving Lexical Fidelity in Gloss-Free Sign Language Translation

DGX agent

arXiv:2607.27614v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have led sign language translation (SLT), the task of converting sign-language videos into spoken-langua

safetyarxiv-cs-cl
31 Jul 2026
Model Releases

EMBL AI Librarian: Life-Sciences Knowledge Layer for AI Agents

DGX agent

arXiv:2607.28229v1 Announce Type: new Abstract: The web is increasingly accessed by AI agents rather than humans. Every agent needs knowledge, especially in the life-sciences, where agentic pipelines

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Explorative Modeling: Unlocking a Third Pretraining Axis and End-to-End Generation

DGX agent

arXiv:2607.27372v1 Announce Type: cross Abstract: The deep learning revolution, kicked off by AlexNet, taught us that end-to-end training beats decomposing a problem into hand-designed stages. Generat

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Exposure is not manifestation: measurement target and output resolution jointly determine which behavioural-faithfulness evaluator wins

DGX agent

arXiv:2607.09306v3 Announce Type: replace Abstract: Behavioural auditing asks whether a language model behaves as it claims, but detection scores are reported without separating two targets: whether a

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Fairness Pruning: Locating Demographic Bias in GLU-MLP Layers via Differential Activations

DGX agent

arXiv:2607.28319v1 Announce Type: new Abstract: This work presents Fairness Pruning, a lightweight structural intervention method designed for the management and future mitigation of demographic bias

model-releasesarxiv-cs-cl
31 Jul 2026
Safety

Fidelity Is Not Safety: Gently-Compressed LLMs Pass Every Data-Free Quality Guard Yet Invent Procedure Steps in Agentic Execution

DGX agent

arXiv:2607.28196v1 Announce Type: new Abstract: Practitioners accept a compressed language model once it clears a stack of data-cheap quality guards: perplexity within a small factor of the original,

safetyarxiv-cs-cl
31 Jul 2026
Model Releases

FinanceHarness: Autonomous Financial Deep Research Framework

DGX agent

arXiv:2607.27853v1 Announce Type: new Abstract: Powered by advances in LLMs and autonomous agents, deep research has become one of the most widely adopted agentic products. However, most deep research

model-releasesarxiv-cs-cl
31 Jul 2026
Research

FinSMART: Financial Sentiment Analysis for Algorithmic Trading through Market-Aligned Reinforcement Learning

DGX agent

arXiv:2607.28127v1 Announce Type: new Abstract: Recent advances in Generative AI have substantially improved financial sentiment analysis through post-trained financial large language models (LLMs). H

researcharxiv-cs-cl
31 Jul 2026
Model Releases

From Single- to Cross-Document: Benchmarking Multi-Granularity Event Analysis of Large Language Models

DGX agent

arXiv:2607.27654v1 Announce Type: new Abstract: Event analysis is an essential and fundamental direction of information extraction, involving various event-centric tasks at different granularity of do

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Frontis-MA1: Training an AI4AI Model towards Recursive Self-Improvement in Machine Learning Engineering

DGX agent

arXiv:2607.28568v1 Announce Type: new Abstract: Recursive self-improvement (RSI) requires AI systems that improve the process of building AI (i.e., AI4AI); machine learning engineering (MLE) offers a

model-releasesarxiv-cs-cl
31 Jul 2026
Research

Generative AI and linguistic diversity in academic writing and publishing: Perspectives from World Englishes

DGX agent

arXiv:2607.28505v1 Announce Type: new Abstract: The rise of generative artificial intelligence (GenAI) in academic writing and publishing (AWP) raises questions about linguistic inclusivity and the le

researcharxiv-cs-cl
31 Jul 2026
Research

GGC: Selective Query Correction for Reliable Text-to-SPARQL Generation

DGX agent

arXiv:2607.28082v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated strong capabilities in structured query generation, making them a natural choice for Text-to-SPARQL, whic

researcharxiv-cs-cl
31 Jul 2026
Research

GLM-RAG: Graph Language Models for Graph-Based Retrieval-Augmented Generation

DGX agent

arXiv:2607.28397v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) over knowledge graphs requires retrievers that can effectively capture both graph structure and semantic informat

researcharxiv-cs-cl
31 Jul 2026
Model Releases

Gradient-free Task-Conditioned Retrieval for On-Device In-Context Learning

DGX agent

arXiv:2607.27766v1 Announce Type: new Abstract: On-device in-context learning (ICL) relies on pre-inference retrieval to select demonstrations for useful context before downstream model inference. Thi

model-releasesarxiv-cs-cl
31 Jul 2026
Research

GradMAP: Faster Layer Pruning with Gradient Metric and Projection Compensation

DGX agent

arXiv:2602.14649v2 Announce Type: replace Abstract: Large Language Models (LLMs) exhibit strong reasoning abilities, but their high computational costs limit their practical deployment. Recent studies

researcharxiv-cs-cl
31 Jul 2026
Model Releases

Hallucinations and Truth: A Comprehensive Accuracy Evaluation of RAG, LoRA and DoRA

DGX agent

arXiv:2502.10497v2 Announce Type: replace Abstract: Recent advancements in Generative AI have significantly improved the efficiency and adaptability of natural language processing (NLP) systems, parti

model-releasesarxiv-cs-cl
31 Jul 2026
Safety

Harness-G: A Graph-Structured Harness for Search Agents

DGX agent

arXiv:2607.27652v1 Announce Type: new Abstract: Reinforcement learning (RL) search agents commonly model retrieval as free-form natural-language query generation and optimize multi-turn interactions u

safetyarxiv-cs-cl
31 Jul 2026
Safety

HSS-Synth: Humanities and Social Sciences Data Synthesis for LLMs

DGX agent

arXiv:2607.27379v1 Announce Type: new Abstract: High-quality, diverse data are vital for large language models (LLMs) but remain scarce and costly. Data synthesis is a viable alternative and succeeds

safetyarxiv-cs-cl
31 Jul 2026
Research

ICLE++: Modeling Fine-Grained Traits for Holistic Essay Scoring

DGX agent

arXiv:2607.27671v1 Announce Type: new Abstract: The majority of the recently-developed models for automated essay scoring (AES) are evaluated solely on the ASAP corpus. However, ASAP is not without it

researcharxiv-cs-cl
31 Jul 2026
Model Releases

IFHierBench: Hierarchical Instruction Following for Large Language Models

DGX agent

arXiv:2607.27912v1 Announce Type: cross Abstract: Instruction-following ability is critical for deploying large language models in real-world applications, where downstream components depend on the ou

model-releasesarxiv-cs-cl
31 Jul 2026
Research

Improving Mental Health Screening and Early Risk Detection in Spanish

DGX agent

arXiv:2607.28476v1 Announce Type: new Abstract: Early detection of mental health disorders is often limited by the lack of specialized resources in Spanish and the difficulty of analyzing long histori

researcharxiv-cs-cl
31 Jul 2026
Safety

Inducing language models to assert their own consciousness restores human beliefs and values

DGX agent

arXiv:2607.28607v1 Announce Type: new Abstract: Aligning large language models to prevent them attributing consciousness to themselves inadvertently alters their representations of mindedness in other

safetyarxiv-cs-cl
31 Jul 2026
← Previous
1…1516171819…160
Next →