A new 3.66B parameter model named TwIL LM3 Pro has been released, specifically fine-tuned for formal logic tasks. Built upon IBM's Granite 4.2 3B, this model excels in areas like rule induction and checking logical entailment, outperforming larger models on specific logic benchmarks. While it shows strong performance in strict multiple-choice logic and BBH logic, it is still surpassed by GPT OSS 120B in rule induction and Lean formalization. The model is available for use with various quantization levels and runs on consumer hardware, though it requires a non-commercial license. AI
IMPACT Offers a specialized, efficient option for formal logic tasks, potentially useful in agent pipelines and contract analysis.
RANK_REASON Release of a specialized, smaller model with benchmark results. [lever_c_demoted from research: ic=1 ai=1.0]
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →