Skip to content
Navigation menu
Search
Powered by Algolia
Search
Log in
Create account
DEV Community
Close
#
llms
Follow
Hide
Posts
Left menu
đź‘‹
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
Right menu
LLMs Reward Expertise: A Shift in AI Agent Design
Felipe L
Felipe L
Felipe L
Follow
Aug 19
LLMs Reward Expertise: A Shift in AI Agent Design
#
llms
#
expertise
#
automation
#
aiagents
Comments
Add Comment
2 min read
Why AI Benchmarks Mean Less Than You Think
The AI Downside
The AI Downside
The AI Downside
Follow
Aug 15
Why AI Benchmarks Mean Less Than You Think
#
benchmarks
#
llms
#
evaluation
#
hype
Comments
Add Comment
6 min read
AI Hallucinations Are Still Not Solved
The AI Downside
The AI Downside
The AI Downside
Follow
Aug 15
AI Hallucinations Are Still Not Solved
#
hallucinations
#
llms
#
reliability
Comments
Add Comment
6 min read
A question about AI I've been carrying for a while
aileen vl
aileen vl
aileen vl
Follow
Jul 18
A question about AI I've been carrying for a while
#
ai
#
llms
#
programming
Comments
Add Comment
5 min read
PoPE: Placebo-Controlled Evaluation Challenges Error-Conditioned Self-Repair in Small Code LLMs
Pneumetron
Pneumetron
Pneumetron
Follow
Jul 15
PoPE: Placebo-Controlled Evaluation Challenges Error-Conditioned Self-Repair in Small Code LLMs
#
llms
#
codegeneration
#
selfrepair
#
evaluation
Comments
Add Comment
3 min read
UniClawBench: A New Benchmark for Proactive AI Agents in Real-World Scenarios
Pneumetron
Pneumetron
Pneumetron
Follow
Jul 13
UniClawBench: A New Benchmark for Proactive AI Agents in Real-World Scenarios
#
aiagents
#
benchmarks
#
llms
#
mllms
Comments
Add Comment
3 min read
IdeaGene-Bench: A New Benchmark for Scientific Lineage Reasoning in AI
Pneumetron
Pneumetron
Pneumetron
Follow
Jul 12
IdeaGene-Bench: A New Benchmark for Scientific Lineage Reasoning in AI
#
aiml
#
benchmarking
#
scientificdiscovery
#
llms
Comments
Add Comment
4 min read
llms.txt: Funktioniert es wirklich? Was Server-Logs zeigen
Tsari Bombelli
Tsari Bombelli
Tsari Bombelli
Follow
Jul 24
llms.txt: Funktioniert es wirklich? Was Server-Logs zeigen
#
llms
#
ai
#
seo
#
devops
Comments
Add Comment
4 min read
How a mesh of peer AI workspaces catches what any single agent misses
David Van Assche (S.L)
David Van Assche (S.L)
David Van Assche (S.L)
Follow
Jul 13
How a mesh of peer AI workspaces catches what any single agent misses
#
ai
#
agents
#
python
#
llms
1
 reaction
Comments
Add Comment
7 min read
The Thinking Inside the LLM Clichés
Arthur
Arthur
Arthur
Follow
Jun 29
The Thinking Inside the LLM Clichés
#
llms
#
writing
#
aiprompts
#
cliches
Comments
Add Comment
7 min read
LLMs: Don't use a sledgehammer when tweezers will do
Louis Van Der Walt
Louis Van Der Walt
Louis Van Der Walt
Follow
Jun 18
LLMs: Don't use a sledgehammer when tweezers will do
#
ai
#
llms
#
dataengineering
Comments
Add Comment
3 min read
Browser-Based Agent Uses WebGPU and Client-Side LLMs to Interact with Web Content Without API Dependencies
Pavel Kostromin
Pavel Kostromin
Pavel Kostromin
Follow
Jun 16
Browser-Based Agent Uses WebGPU and Client-Side LLMs to Interact with Web Content Without API Dependencies
#
webgpu
#
llms
#
privacy
#
autonomy
Comments
Add Comment
14 min read
The Chomsky Objection the AI Industry Has Been Quietly Working Around
Arthur
Arthur
Arthur
Follow
Jun 9
The Chomsky Objection the AI Industry Has Been Quietly Working Around
#
llms
#
aiphilosophy