« All posts

GLM-4.7-Flash on 2x RTX 3090: My Hands-On Experience

GLM-4.7-Flash was tested on 2x RTX 3090. Performance comparison in short and long contexts was conducted.

The GLM-4.7-Flash model was tested on two RTX 3090 cards. In short contexts, GLM showed a significant speed advantage, while in long contexts, the simpler Qwen3.5:27b model outperformed it. This finding is crucial for engineers to determine which model is more efficient in specific scenarios.