LLM Prompt Duel Optimizer: Efficient Label-Free Prompt Optimization
1月 29, 2026admin NUAI,Committee,ニュース,Uncategorized(0)arXiv:2510.13907v2 Announce Type: replace Abstract: Large language models (LLMs) are highly sensitive to prompts, but...
LLM Probing with Contrastive Eigenproblems: Improving Understanding and Applicability of CCS
11月 5, 2025admin NUAI,Committee,ニュース,Uncategorized(0)arXiv:2511.02089v1 Announce Type: cross Abstract: Contrast-Consistent Search (CCS) is an unsupervised probing method able to...
LLM Orchestration Frameworks Compared: LangChain vs. LlamaIndex vs. Raw API Calls
7月 9, 2026admin NUAI,Committee,ニュース,Uncategorized(0)The default assumption in most LLM developer communities is that you start with raw API...

LLM or Human? Perceptions of Trust and Information Quality in Research Summaries
1月 23, 2026admin NUAI,Committee,ニュース,Uncategorized(0)arXiv:2601.15556v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used to generate and...
LLM one-shot style transfer for Authorship Attribution and Verification
10月 16, 2025admin NUAI,Committee,ニュース,Uncategorized(0)arXiv:2510.13302v1 Announce Type: new Abstract: Computational stylometry analyzes writing style through quantitative patterns in text...
LLM Observability Tools for Reliable AI Applications
5月 12, 2026admin NUAI,Committee,ニュース,Uncategorized(0)Large language models (LLMs) now power everything from customer service bots to autonomous coding agents...

LLM Evaluation Frameworks Compared: How to Actually Measure What Your Model Does
7月 14, 2026admin NUAI,Committee,ニュース,Uncategorized(0)In this article, you will learn how to evaluate LLM applications using the three dominant...

LLM Embeddings vs TF-IDF vs Bag-of-Words: Which Works Better in Scikit-learn?
2月 17, 2026admin NUAI,Committee,ニュース,Uncategorized(0)Machine learning models built with frameworks like scikit-learn can accommodate unstructured data like text, as...
LLM BiasScope: A Real-Time Bias Analysis Platform for Comparative LLM Evaluation
3月 16, 2026admin NUAI,Committee,ニュース,Uncategorized(0)arXiv:2603.12522v1 Announce Type: new Abstract: As large language models (LLMs) are deployed widely, detecting and...