Explore Prompts

Page 326 of 511 · 6130 prompts

Multilingual Alignment @ Scale

Design a multilingual alignment plan (zh/ja/hi/id/pt/en): shared subword policy, cross-lingual instructions, and locale-specific refusal tuning. Provide leakage checks.
Tags: LLM, multilingual, alignment, tokenization, refusal-tuning
Author: Assistant
Created at: 2025-12-18 00:00:00
Average Rating:
Total Ratings:

Multi-Task Multi-Domain Evals

Create a senior-grade eval battery: reasoning (math/code), instruction-following, safety, multilingual QA, and tool-use. Include uncertainty intervals and power analysis for A/Bs.
Tags: LLM, evaluation, multidomain, statistics, AB-testing
Author: Assistant
Created at: 2025-12-18 00:00:00
Average Rating:
Total Ratings:

Structured Output Contracts

Define JSON schema contracts with type coercion, partial output recovery, and EBNF constraints. Provide test-time correction and repair strategies.
Tags: LLM, structured-output, JSON, EBNF, validation, repair
Author: Assistant
Created at: 2025-12-18 00:00:00
Average Rating:
Total Ratings:

Agents: Planner–Executor–Critic

Specify a lightweight agent loop with decomposition, execution, and critique. Provide termination conditions, trace logging, and loop unroll limits.
Tags: LLM, agents, planning, critique, traces, governance
Author: Assistant
Created at: 2025-12-18 00:00:00
Average Rating:
Total Ratings:

Toolformer-Style Tool Use

Design a tool-use curriculum: function signatures, schema discovery, tool reliability scoring, and retry/backoff policy. Include sandboxing and cost guards.
Tags: LLM, tools, function-calling, reliability, sandbox, cost
Author: Assistant
Created at: 2025-12-18 00:00:00
Average Rating:
Total Ratings:

Hallucination Detection & Abstain

Create a hallucination detector using entailment+attribution signals. Define abstention thresholds, user messaging, and a re-query strategy with targeted retrieval.
Tags: LLM, hallucination, entailment, abstention, UX, grounding
Author: Assistant
Created at: 2025-12-18 00:00:00
Average Rating:
Total Ratings:

Retrieval Eval Harness

Build an eval harness: recall@k, calibrated precision, answer faithfulness, and human-time-to-verify. Include topic-aware test buckets and data drift alarms.
Tags: LLM, retrieval, eval, faithfulness, drift, metrics
Author: Assistant
Created at: 2025-12-18 00:00:00
Average Rating:
Total Ratings:

RAG 2.0: Freshness & Faithfulness

Architect a retrieval stack with hybrid search, temporal decay, dedup, and passage-level citation anchors. Define fact-grounding checks and failure messages; include freshness reindex cadence.
Tags: LLM, RAG, hybrid, temporal-decay, citations, freshness
Author: Assistant
Created at: 2025-12-18 00:00:00
Average Rating:
Total Ratings:

Synthetic Data: Self-Play & Critique

Propose a self-play generation strategy where a teacher model drafts, a critic model scores, and a curator enforces diversity/novelty. Provide leakage and drift monitors.
Tags: LLM, synthetic-data, self-play, critic, curation, drift
Author: Assistant
Created at: 2025-12-18 00:00:00
Average Rating:
Total Ratings:

Instruction Mining @ Scale

Design a pipeline to mine high-quality instructions/solutions from forums, docs, and code. Include classifier-based filtering, self-checking, and multilingual normalization.
Tags: LLM, instruction-mining, classification, self-check, ETL, multilingual
Author: Assistant
Created at: 2025-12-18 00:00:00
Average Rating:
Total Ratings:

Data Governance & Decontamination

Write a data governance spec: license screening, PII scrubbing, near-duplicate collapse, contamination checks vs eval sets, and audit trails. Provide rejection reasons and exception handling.
Tags: LLM, data-governance, PII, decontamination, licensing, audit
Author: Assistant
Created at: 2025-12-18 00:00:00
Average Rating:
Total Ratings:

Fine-Tune Stack: SFT→DPO/ORPO→RLHF

Specify a training stack with SFT on curated data, preference optimization (DPO/ORPO), and optional RLHF. Include reward hacking tests, guardrails, and evals that predict production behavior.
Tags: LLM, SFT, DPO, ORPO, RLHF, alignment, evaluation
Author: Assistant
Created at: 2025-12-18 00:00:00
Average Rating:
Total Ratings: