Skip to main content
PaperintermediateFree

Toolformer: Language Models Can Teach Themselves to Use Tools

by Timo Schick et al.

Presents a self-supervised method by which a language model annotates its own training data with API calls to a calculator, search engine, translator or calendar, then fine-tunes on the calls that reduce perplexity, learning when to invoke external tools.

Visit resource

More resources on AI Agents

CourseFree

Multi Agent Systems

Builds teams of cooperating agents with the crewAI framework, assigning roles, tools, memory, and task decomposition. Worked examples cover resume tailoring, technical article writing, customer support, outreach, event planning, and financial analysis.

CourseFree

Agent Design Patterns

Covers core agentic patterns (reflection, tool use, planning, and multi-agent group chat) implemented in Microsoft's AutoGen framework. Projects include a two-agent conversation, a reflective blog writer, a chess-playing agent, and coding agents for financial analysis.

CourseFree

Multi-Agent Use

Advanced crewAI course focused on shipping agent systems: external integrations, coordinating multiple models in one crew, testing with human feedback, and deployment. Projects include project planning, a Trello progress reporter, a sales pipeline, and support analytics.

CourseFree

Computer Use with Anthropic

Progresses from the Claude API through multimodal prompting, prompt caching, and tool calling to Anthropic's computer use feature, ending with an agent that reads screenshots and operates a desktop interface to complete tasks.

CourseFree

Evaluating AI Agents

Treats agent evaluation as its own discipline: adding tracing and observability, choosing between code-based checks, LLM-as-judge, and human review, then scoring router decisions, individual skills, and full trajectories through structured experiments.

CourseFree

Building Browser Agents

Explains how web agents perceive pages through visual and DOM structure, then plan actions like scraping, summarizing, and form filling. Also covers AgentQ, which combines Monte Carlo tree search, self-critique, and DPO for self-correcting agents.

See all AI Agents resources →