PAN: A World Model for General, Interactable, and Long-Horizon World Simulation Paper • 2511.09057 • Published 25 days ago • 75
FinAuditing: A Financial Taxonomy-Structured Multi-Document Benchmark for Evaluating LLMs Paper • 2510.08886 • Published Oct 10 • 19
FinChain: A Symbolic Benchmark for Verifiable Chain-of-Thought Financial Reasoning Paper • 2506.02515 • Published Jun 3 • 3
LLM-DetectAIve: a Tool for Fine-Grained Machine-Generated Text Detection Paper • 2408.04284 • Published Aug 8, 2024 • 26
Llama 3.2 Collection This collection hosts the transformers and original repos of the Llama 3.2 and Llama Guard 3 • 15 items • Updated Dec 6, 2024 • 647