Mayank Maheshwari
mayankiit04
AI & ML interests
None yet
Recent Activity
new activity about 12 hours ago
unsloth/Qwen3.8-Flash-Next-GGUF:unsloth studio vs strata engine for qwen3.8-flash-next new activity 2 days ago
unsloth/Qwen3.8-Flash-Next-GGUF:will the new llama.cpp build 11330+ support the mtp file?Organizations
None yet
unsloth studio vs strata engine for qwen3.8-flash-next
2
#80 opened 1 day ago
by
mayankiit04
llama.cpp now supports MTP in main, but these quants don't seem compatible
2
#79 opened 2 days ago
by
FlorinAndrei
will the new llama.cpp build 11330+ support the mtp file?
π₯π 2
3
#78 opened 2 days ago
by
mayankiit04
unsloth shared MTP not working with this IQ3_S on llama.cpp with build 11330
π 2
1
#36 opened 2 days ago
by
mayankiit04
IQ3_S
π₯ 6
17
#26 opened 8 days ago
by
DzmitryTheOtherOne
What is the SWE Verfied of the the pure GSQ RCO non coder Q3_S
1
#2 opened 7 days ago
by
mayankiit04
Quick share: Got 30β40 t/s on the RTX 5060 Ti 16GB
12
#61 opened about 1 month ago
by
QilinWan
Small gain with MTP
7
#59 opened about 1 month ago
by
OmarColocci
can we have a gguf varity where ngram layer 2 is at q8?
#53 opened about 1 month ago
by
mayankiit04
How can I improve the prefilling speed for this model?
14
#49 opened about 1 month ago
by
BipedalBit
With Qwen and GLM now in sub 300B local machine level, this is a big model.
π 1
1
#10 opened about 1 month ago
by
mayankiit04
Is the n-gram chunk embedded in the gguf(s), and is it ~51GB independent of quantization?
14
#35 opened about 1 month ago
by
dagb
How to keep n-gram table on fast nvme ssd
24
#23 opened about 1 month ago
by
mayankiit04
Wow this is an outright killer model... anyone can now run this with mere 48 gb ram/vram & a fast nvme drive
6
#30 opened about 1 month ago
by
mayankiit04
here qwen3.8-flash-next is 125B but unsloth has 180B ... how come?
3
#17 opened about 1 month ago
by
mayankiit04
Are you serious, Chatgpt puts this model at estimate 59 score AA above opus 4.8!!
π 1
#3 opened about 1 month ago
by
mayankiit04
ArtificialAnalysis score of 52 outscore GLM 5.2 , Opus 4.6 and touches Opus 4.7!!!!
π₯ 1
10
#143 opened about 2 months ago
by
mayankiit04