You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You switched accounts on another tab or window. Reload to refresh your session.Dismiss alert
Our current center of gravity is evidence-grounded output verification: given source material and an answer, locate unsupported, contradictory, or fabricated spans. Code answers and tool-based agent responses appear to fit that model; content safety and jailbreak detection may be adjacent rather than the same product.
Which workloads should belong in LettuceDetect? Please describe:
the source evidence;
the generated output;
the failure you need localized;
why an answer-level score is insufficient.
Concrete counterexamples and reasons a workload does not fit are especially useful.
reacted with thumbs up emoji reacted with thumbs down emoji reacted with laugh emoji reacted with hooray emoji reacted with confused emoji reacted with heart emoji reacted with rocket emoji reacted with eyes emoji
Uh oh!
There was an error while loading. Please reload this page.
Our current center of gravity is evidence-grounded output verification: given source material and an answer, locate unsupported, contradictory, or fabricated spans. Code answers and tool-based agent responses appear to fit that model; content safety and jailbreak detection may be adjacent rather than the same product.
Which workloads should belong in LettuceDetect? Please describe:
Concrete counterexamples and reasons a workload does not fit are especially useful.
All reactions