TechJun 8, 202617 minLFM2.5 1.2B JP on M1 Max 64GB: 208 tok/s decode, JSON OK, name hallucinatedTested LFM2.5-1.2B-JP-202606 on M1 Max 64GB. llama.cpp Q4_K_M: 208 tok/s decode, JSON intact, model name hallucinated (LFM→FDM). Q8_0: 157 tok/s, no hallucination. Tool calls broken via GGUF.AILLMLocal LLMMLXOllamaApple SiliconEdge AIExperimentJapanese LLM