RAND
Popular repositories Loading
-
judge-reliability-harness
judge-reliability-harness PublicFramework for generating validation data for LLM-as-a-Judge and LLM Autograders
-
milliondigits
milliondigits PublicIn 1955, RAND Corporation published a book of a “million random digits”; this code is a modern attempt to reproduce the analysis that was originally conducted on punchcards. This new analysis does …
Repositories
- achieving-ai-model-weight-sl3 Public
Supporting materials for the RAND report "Achieving AI Model Weight Security Level 3" (RR-A4704-1)
-
- LibreChat Public Forked from danny-avila/LibreChat
Enhanced ChatGPT Clone: Features Agents, MCP, DeepSeek, Anthropic, AWS, OpenAI, Responses API, Azure, Groq, o1, GPT-5, Mistral, OpenRouter, Vertex AI, Gemini, Artifacts, AI model switching, message search, Code Interpreter, langchain, DALL-E-3, OpenAPI Actions, Functions, Secure Multi-User Auth, Presets, open-source for self-hosting. Active.
- mh-llm-eval-public Public
Replication code and data for Suicide-Related Responses from AI Chatbots through Consumer-Facing Interfaces and APIs (Cantor et al., 2026).
- missile-jamming-and-engagement-model Public
Python implementation of a model that estimates measures of performance and effectiveness for air-launched missile engagements under GPS and IFTU jamming conditions
- autograders-rr-2026 Public
Simpler is Better for AI Autograders: Towards Cost-Effective LLM Evaluations For Open-Ended Tasks
- mced-spillover-2025 Public
Code and data used for analysis described in manuscript 'Potential spillover effects on diagnostic delay for cancer during the NHS-Galleri trial: a quasi-experimental difference-in-differences study
Top languages
Loading…
Most used topics
Loading…