Skip to contentSkip to stories

Updated

All AI news

Apr 7

Apr 7Tue

Apr 5

Apr 5Sun

Apr 4

Apr 4Sat

Apr 2

Apr 2Thu
  1. AI Futures ProjectAI score62

    AI Futures Project shortens Automated Coder timelines to mid 2028

    AIAI Futures Project moved Daniel Kokotajlo's Automated Coder median from late 2029 to mid 2028 and Eli's from early 2032 to mid 2030. The main reasons cited are a faster METR time horizon doubling time and the impressive results of Claude Opus 4.6. The authors also say progress in agentic coding has been faster than expected over the past 3 to 5 months.

Apr 1

Apr 1Wed

Mar 30

Mar 30Mon

Mar 28

Mar 28Sat
  1. Andrej KarpathyAI score12

    Karpathy: LLMs can argue both sides, so beware sycophancy

    AIAndrej Karpathy reports that an LLM spent four hours strengthening his blog post's argument, then convinced him of the opposite when asked to argue the reverse. He concludes that LLMs are highly capable of arguing almost any direction, which makes them useful for forming opinions if users ask from multiple angles and watch for sycophancy.

Mar 26

Mar 26Thu
  1. Andrej KarpathyAI score47

    Karpathy wants agents to handle full app DevOps from one command

    AIAndrej Karpathy argues that the hardest part of building a deployed app is not the code but the DevOps work of assembling services, API keys, payments, auth, and deployment. He says the goal is for agents to handle this entire lifecycle as code, with agent-native CLI and API access instead of manual web clicking. He calls it a from-scratch redesign that is only now barely technically possible.

  2. Hamel HusainAI score38

    Data Scientists Face New Pressures as LLM APIs Let Teams Ship AI Without Them

    AIHamel Husain argues data scientists remain essential as foundation-model APIs let teams ship AI without them, because much of the work lies in evaluation, debugging, and metric design. He says teams often rely on generic off-the-shelf metrics and unverified LLM judges instead of examining their own data. He lists five eval pitfalls, starting with generic metrics, and recommends looking at traces and doing error analysis.

Mar 25

Mar 25Wed

Mar 24

Mar 24Tue
  1. Jim FanAI score62

    Jim Fan warns that compromised LiteLLM package shows risks for AI agents

    AIJim Fan reposted a report that LiteLLM PyPI release 1.82.8 was compromised and contained a litellm_init.pth file that sends credentials to a remote server and self-replicates. He argues agents make this worse, since files like skills, configs, or PDFs read into context could spread malicious instructions. He concludes that agentic frameworks need guardrails and audited tooling.

Mar 23

Mar 23Mon

Mar 19

Mar 19Thu

Mar 18

Mar 18Wed

Mar 13

Mar 13Fri

Mar 1

Mar 1Sun
  1. Chris OlahAI score62

    Legal analyst says OpenAI's Pentagon contract language only guarantees all lawful use

    AIThe author shares a quoted legal analysis arguing that OpenAI's published Pentagon contract excerpt essentially only permits all lawful use. The analyst notes the excerpt is short, that DoD Directive 3000.09 and other DoD directives referenced in it can be changed by the Department at any time, and that the contract may not guarantee what OpenAI's FAQ implies.

Feb 28

Feb 28Sat

Feb 27

Feb 27Fri

Feb 26

Feb 26Thu

Feb 25

Feb 25Wed

Feb 24

Feb 24Tue

Feb 23

Feb 23Mon