Non-Visual AI Productivity

AI workflows designed to work beautifully with screen readers, keyboard-only navigation, and clear structure.

Back to Lessons

Lesson 5: Multimodal AI and Agents (Non-Visual Productivity)

Course: Foundations of Non-Visual AI Productivity (AI Basics)

Lesson content

Continue

All lessons in this course

References

  1. 1. OpenAI, “GPT-4 Technical Report” (2023). https://arxiv.org/abs/2303.08774
  2. 2. OpenAI, “GPT-4V(ision) System Card” (2023). https://cdn.openai.com/papers/GPTV_System_Card.pdf
  3. 3. Yao et al., “ReAct: Synergizing Reasoning and Acting in Language Models” (2022). https://arxiv.org/abs/2210.03629
  4. 4. Park et al., “Generative Agents: Interactive Simulacra of Human Behavior” (2023). https://arxiv.org/abs/2304.03442
  5. 5. Stanford HAI, “AI Index Report” (latest edition). https://aiindex.stanford.edu/report/