Notizie
Notizie
Prompt-R1: Collaborative Automatic Prompting Framework via End-to-end Reinforcement Learning
arXiv:2511.01016v3 Announce Type: replace Abstract: Recently, advanced large language models (LLMs) have emerged at an...
Prompt Optimization as a State-Space Search Problem
arXiv:2511.18619v1 Announce Type: new Abstract: Language Models are extremely susceptible to performance collapse with even...
Prompt Engineering vs Loop Engineering vs Graph Engineering: What Changes at Each Layer
Three terms now compete for the same line in AI engineering job descriptions. Prompt engineering...
Prompt Engineering for Agentic AI
You have probably spent time learning how to prompt AI well...
Prompt engineering does not universally improve Large Language Model performance across clinical decision-making tasks
arXiv:2512.22966v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated promise in medical knowledge...
Procedural Knowledge at Scale Improves Reasoning
arXiv:2604.01348v3 Announce Type: replace Abstract: Test-time scaling has emerged as an effective way to improve...
Probing the Critical Point (CritPt) of AI Reasoning: a Frontier Physics Research Benchmark
arXiv:2509.26574v3 Announce Type: replace-cross Abstract: While large language models (LLMs) with reasoning capabilities are progressing...
Probing Neural Topology of Large Language Models
arXiv:2506.01042v3 Announce Type: replace Abstract: Probing large language models (LLMs) has yielded valuable insights into...
Probabilistic distances-based hallucination detection in LLMs with RAG
arXiv:2506.09886v2 Announce Type: replace Abstract: Detecting hallucinations in large language models (LLMs) is critical for...
Probabilistic Aggregation and Targeted Embedding Optimization for Collective Moral Reasoning in Large Language Models
arXiv:2506.14625v2 Announce Type: replace Abstract: Large Language Models (LLMs) have shown impressive moral reasoning abilities...
Proactive Defense: Compound AI for Detecting Persuasion Attacks and Measuring Inoculation Effectiveness
arXiv:2511.21749v1 Announce Type: new Abstract: This paper introduces BRIES, a novel compound AI architecture designed...
Proactive defense against LLM Jailbreak
arXiv:2510.05052v2 Announce Type: replace-cross Abstract: The proliferation of powerful large language models (LLMs) has necessitated...
