Llm
Topic archive • 6 matches
2026-09-14
Technology
Managed Agents - Don't Get Locked In: This video analyzes the concept of managed agents currently being rolled out by various Large Language Model providers. While these managed services offer numerous advantages, they also present a significant risk of vendor lock-in for developers.
LLM Agents • Sam Witteveen
Permalink
Tips
Large Language Models
Audit identifies 12 data leaks and compliance risks in agentic LLM pipelines
Permalink
2026-09-12
Technology
LogiMed-RoB benchmark reveals error compounding in LLM medical logic: Researchers have introduced LogiMed-RoB, a benchmark based on Cochrane Risk of Bias 2.0 expert logic to evaluate large language models across 860 randomized controlled trials. Testing on 10 state-of-the-art models revealed a severe error compounding effect, despite the top model achieving 98.88% atomic consistency.
AI Models and Applications • arXiv
Permalink
2026-09-07
Technology
More capable LLMs can increase systemic risk in financial markets: A study shows that improving individual large language model capability can degrade system-level outcomes rather than improve them. Researchers hypothesize that shared training and architectures cause more capable LLMs to behave similarly, creating correlated actions that do not diversify away.
Research • arXiv
Permalink