SkinDeepRESEARCHSteve Seguin

Early tests

Decisions without a written reply

Sometimes an app only needs a label, such as “support” or “spam”.

Where is my order?Support
Illustration: route a message without writing a reply.

An app may need a category or a yes/no decision. A classifier can return that result directly, without writing a sentence or JSON.

This changes the output. The model can still use every layer; stopping early is a separate optimization.

Run 100 real moderation messages in your browser · Watch recorded movement decisions

Try it on your device

The browser benchmark runs Qwen 0.5B on real ToxicChat messages. Compare a direct label decision with a written JSON answer, including mistakes and the time to finish. Both use all 24 layers.

Where this could be useful

When an app needs a category, score or action, a small output head can return it directly. Some cases have narrow prototypes; others are proposals.

More use cases: search, actions, signals and image checks

Messages and documents

Actions

Medical signals and scans

Language and image generation

What the research has shown

A separately trained Qwen classifier got 2,551 of 3,080 banking queries right across 77 topics. A word-based classifier got 2,464 right. The browser demo uses Qwen’s existing output scores; it does not reproduce those trained classifiers.

A separate changing-rule test failed. Following arbitrary instructions remains unproven, and shorter output does not guarantee equal accuracy.

Banking data and results · Matched one-token comparison · Changing-rule test · Other ways to reduce computation

Five more practical datasets to test

Processing several messages together

Reusing the fixed instructions: measured results

Try a specific task

A decision on every step