Wang
Liuwang971
AI & ML interests
None yet
Organizations
None yet
[fastest inference] locateanything-batch β batched + KV-cached LocateAnything-3B, ~2.7Γ faster
ππ 6
1
#10 opened 2 months ago
by
Liuwang971
Benchmark: locateanything-batchοΌMTPοΌ vs the llama.cpp & vLLM AR ports
π 2
2
#14 opened 2 months ago
by
Liuwang971
Inference support for vLLM and SGLang OpenAI endpoints
β 14
10
#3 opened 2 months ago
by
Vishva007
Batch query
2
#5 opened 2 months ago
by
SRai22