GLM-4-Flash is a free lightweight text LLM developed by Zhipu AI. It features a 128K context window and supports function calling, multilingual tasks and web search. With fast inference, it fits high-concurrency scenarios such as text processing, Q&A and summarization.