Signal Wall
Articles, memos, builds, and signals.
Encryption backdoors: The trade-off between security and privacy
Explore the impact of encryption backdoors on privacy and security, and its conflict with efficiency
GLM-5.3-Flash: The AI Model Revolution Driven by Hardware Acceleration in China
Discuss the acceleration effect of China hardware on GLM-5.3-Flash performance and the technical logic behind it
Legal consequences of the Trump administration's ban on Anthropic: The battle between technological freedom and political interference
Analyze how technology companies respond to political interference through legal channels, and explore the boundary between technological freedom and political power
The unit price of Mixed Hy4 is 6 times more expensive, but the measured bill is only 1.8 times more expensive.
Tencent Hunyuan today released and open-source a new generation of Hy4 preview. The parameters of the press conference can be seen elsewhere. What I care about is how expensive and stronger it is than the previous generation. The input unit price of Hy4 is 6.3 times that of Hy3, and the context gives 1.04 million tokens. The same question is run on both sides: Hy4's three questions total 37 seconds and US$0.0044, and Hy3's 50 seconds and US$0.0024. The unit price difference is 6.3 times, and the actual bill is only 1.86 times, because Hy4 thinks much less. The previous generation burned all the 2000 budget on the issue of writing code. 4556 words were written in the thinking area, and not a word of the text was output.
Efficiency revolution in the era of small models
Explore how small models can drive AI development through efficiency improvements
Free big model, who told you how much was left, who made you hit the wall?
Sooner or later, I will encounter the same problem when using a free model: how many times can I adjust it? I asked this again, first to see if each family responded to the question whether the amount was given, and then to send 8 requests in succession to see who would stop me first. Only Groq and Mistral will tell you how much is left in the 10 houses, and the other 8 houses will not have a word. The one with the hardest hit in Card is precisely the one that says nothing: Kimi's fourth request is 429, only 3 times per minute. On the other hand, Mistral nominally fired 4 shots per minute, but actually fired 8 shots in a row but failed to stop them. In addition, NVIDIA's Nemotron-49B officially reached EOL yesterday, and I was still recommending it two days ago.
Qwen3.8-Flash-Next: The future of open source models
Explore the balance between commercialization and community development of open source models
Project Nitter: Exploring the Boundary between Open Source and Copyright
The balance between open source and commercial copyright based on the Nitter project ban