Hacker News
new
|
past
|
comments
|
ask
|
show
|
jobs
|
submit
|
electroglyph's comments
login
electroglyph
1 day ago
|
parent
|
context
|
next
[–]
| on:
Show HN: Cactus Needle 3: 8-29MB automation models...
those are all expecting far too much for models this size
reply
p1necone
1 day ago
|
parent
|
next
[–]
Given the title of the post says 'can match deepseek v4 flash' I think it's fair to call out these sort of dumb mistakes.
reply
electroglyph
1 day ago
|
parent
|
context
|
prev
|
next
[–]
| on:
Claude Code now reads AGENTS.md if there is no Cla...
do you just manually type out your important instructions every time instead of being smart and putting a few lines in a text file?
reply
electroglyph
1 day ago
|
parent
|
context
|
prev
|
next
[–]
| on:
Shapelearn Qwen 3.8 27B (13.1 GB VRAM)
prismml's title is very misleading. in their own paper the model is at 75% of coding scores.
reply
electroglyph
20 days ago
|
parent
|
context
|
prev
|
next
[–]
| on:
How to build a diffusion language model
good stuff, no mention of confidence tho, recommend having a look at diffusiongemma and others.
electroglyph
26 days ago
|
parent
|
context
|
prev
|
next
[–]
| on:
Anthropic's best AI model struggles to attract use...
Opus 5 literally benchmarks higher than Fable
electroglyph
28 days ago
|
parent
|
context
|
prev
|
next
[–]
| on:
One night in Uzbekistan: Why was this one data poi...
Good on them for the retraction. It's good to see science at work.
electroglyph
34 days ago
|
parent
|
context
|
prev
|
next
[–]
| on:
An image can overflow
it's not a schooner...it's a sailboat
electroglyph
39 days ago
|
parent
|
context
|
prev
|
next
[–]
| on:
Show HN: Needle2: 14MB agentic LLM for phones, wea...
i would assume a model this size would require finetuning tbh. even functiongemma recommends that.
electroglyph
43 days ago
|
parent
|
context
|
prev
|
next
[–]
| on:
DeepSeek V4 Flash 0731
my experience is the same, but deepseek is planning on increasing prices soon, which will make it a lot less attractive
electroglyph
43 days ago
|
parent
|
context
|
prev
|
next
[–]
| on:
Improving GPT‑5.6 Sol in ChatGPT, expanding GPT‑5....
yep, agreed. little bit smarter than deepseek 4 flash, but a little bit worse at tool calls.
More
Guidelines
|
FAQ
|
Lists
|
API
|
Security
|
Legal
|
Apply to YC
|
Contact
Search:
reply