Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
๐๏ธ
Building on HF
Fontlab Ltd
PRO
fontlab
1
218
Follow
21world's profile picture
Quazim0t0's profile picture
Banaxi-Tech's profile picture
4 followers
ยท
11 following
https://www.fontlab.com/
fontlab
fontlabltd
AI & ML interests
Fonts, vector graphics, multilingual text processing
Recent Activity
reacted
to
OppaAI
's
post
with ๐
about 5 hours ago
Benchmark test: Jev vs. Laya-ONNX (multilingual) vs. Harrier OSS 270M embedder ๐ฌ My AI wAIfu (Jetson Orin Nano 8GB) uses Harrier OSS 270M for semantic routing in 2 places. It reads vectors of router prompts (English only) and calculates cosine similarity: - Quaternary routing: greeting, local chat (no websearch), web chat (needs websearch), or agentic chat - Agentic routing: which tools in my AI's capability list to use Benchmarked the 2 most hyped decision models โ Jev and Laya (ONNX, multilingual) โ against Harrier OSS 270M. Setup: 221 quaternary + 58 capability-trigger examples, leave-one-out eval, argmax, no thresholds. Results: โ Harrier-270M (local, cosine): 94.6% / 93.1% accuracy, 17ms P50 โก โ Jev API (hosted): 82.4% / 94.8% accuracy, ~195ms P50 โ Laya-ONNX multilingual (fp16, local): 48.0% / 20.7% accuracy, 25-40ms P50 Conclusion: ๐ซ Laya is out of the question. 4 of 7 capability categories at 0.0% accuracy while reporting 80-90% confidence means it needs real training before it's practical. โ๏ธ Jev is a cloud API, not sure it can be trained further. Accuracy is high but not improvable on my end. Latency is ~10x my local embedder (network latency). Input token cost, though small, is still more than $0. Not fully sure about privacy implications either. โ Embedding is only semantic cosine similarity, not real reasoning. But it's already doing double duty for memory extraction and RAG โ no extra RAM or token cost. Latency is 17ms, accuracy in the 90s%. Even tried Japanese/Chinese prompts, still got high accuracy with only English exemplars. Bigger advantage: I just add exemplars to boost accuracy. When I add/modify/remove tools โ often โ no retraining needed, just update exemplars, vectors recompute once. Turns out my self-invented routing method, built ~6 months ago, already solved what these now hyped up models โ beating Jev and Laya on latency and convenience, matching/beating on accuracy. ๐ฏ
liked
a model
about 6 hours ago
n4ze3m/Qwen3.5-4B-Hmm
liked
a model
2 days ago
mradermacher/decider-0.8b-GGUF
View all activity
Organizations
models
7
Sort:ย Recently updated
fontlab/BananaMind-2-Mini-Chat-int8
Text Generation
โข
25.6M
โข
Updated
Aug 17
โข
43
โข
1
fontlab/BananaMind-2-Pro-Preview-Chat-ternary
Text Generation
โข
54.5M
โข
Updated
Aug 17
โข
59
fontlab/BananaMind-2-Nano-Chat-ternary
Text Generation
โข
4.35M
โข
Updated
Aug 17
โข
65
fontlab/BananaMind-2-Pro-Preview-Chat-mixed
Text Generation
โข
0.1B
โข
Updated
Aug 17
โข
55
fontlab/BananaMind-2-Pro-Preview-Chat-int8
Text Generation
โข
0.1B
โข
Updated
Aug 17
โข
20
fontlab/BananaMind-2-Nano-Chat-mixed
Text Generation
โข
9.98M
โข
Updated
Aug 17
โข
64
fontlab/BananaMind-2-Nano-Chat-int8
Text Generation
โข
10.1M
โข
Updated
Aug 17
โข
22
datasets
0
None public yet