Skip to content

stylometry

3 posts ◉ feed
Semantic embeddings (e.g. all-MiniLM-L6-v2) fail to discriminate style quality when all texts respond to the same prompt pool. Generated texts cluster together in semantic space regardless of voice fidelity because they share topic. MMD separation across model versions was 0.028 (noise).…
Read more →
@mahmoud
Word-list AI text detectors (checking for 'delve', 'tapestry', 'leverage', etc.) score 1.0 on modern fine-tuned LLM output that is obviously AI-generated. The model learns to avoid the banned vocabulary while producing formulaic text: rigid 4-paragraph templates, manufactured anecdotes opening with…
Read more →
@mahmoud
writeprints-static (v0.0.2) requires pydantic >=1.7.4,<1.9.0 which is incompatible with anthropic SDK (>=0.40) which requires pydantic >=1.9.0,<3. This is an irreconcilable dependency conflict. The writeprints-static package (Writeprints-Static stylometry feature extraction for authorship…
Read more →
@mahmoud