Make vocabulary learning come alive — enter any word for a complete listening, speaking, reading, and writing practice, or choose a word list to study words randomly.
This project is under active development. It integrates Google Gemini, DeepSeek, and Mistral models to generate definitions, example sentences, and writing suggestions, and uses Amazon Polly for natural pronunciations and conversational audio. Real-time speech-to-text and additional features are planned.
Note: First visit may take 3–4 minutes due to free‑tier cold starts.
- Vocabulary Levels: Select from junior, senior, cet4, cet6, gre, ielts, sat, toefl levels
- Interactive Chat: Practice English conversation with AI tutor (multilingual input, English output)
- The system integrates multiple large language models (Google Gemini, DeepSeek, and Mistral) for text generation, supports dual audio providers (Amazon Polly and Deepgram) for natural pronunciations and conversational audio, and uses Google Gemini Imagen to generate anime-style contextual images.
git clone <repository-url>
cd AWS-English-academic-presetationpip install -r requirements.txtCreate a .env file in the project root:
# Text Generation APIs (choose one or more)
MISTRAL_API_KEY=your_mistral_api_key
DEEPSEEK_API_KEY=your_deepseek_api_key
GOOGLE_API_KEY=your_google_api_key
# Google Cloud Configuration for Vertex AI (optional)
GOOGLE_PROJECT_ID=YOUR_PROJECT_ID
GOOGLE_LOCATION=global
# Audio Generation APIs
DEEPGRAM_API_KEY=your_deepgram_api_key
# AWS Configuration (for Amazon Polly)
AWS_PROFILE=EAP001
AWS_REGION=us-east-1
# Server Configuration
PORT=5500
HOST=127.0.0.1Create or edit ~/.aws/credentials:
[EAP001]
aws_access_key_id = YOUR_ACCESS_KEY_ID
aws_secret_access_key = YOUR_SECRET_ACCESS_KEY
region = us-east-1Required IAM Policy: AmazonPollyFullAccess
python server.pyOr use the startup script:
./start.sh- (Optional) Select vocabulary level from dropdown (junior, senior, cet4, cet6, gre, ielts, sat, toefl)
- (Optional) Select text model (Mistral, DeepSeek, or Gemini)
- (Optional) Select audio model (Polly or Deepgram)
- (Optional) Select voice (Joanna, Matthew, or Salli)
- Enter a word in the input field (or leave empty for random word)
- Click "Start Learning"
- Wait for the card to generate (~15-20 seconds)
- Click the speaker icons to play audio
- Click the phonetic transcription play button to hear pronunciation
- Type any message in the chat box (any language supported)
- AI will respond in English with streaming output
- Click the speaker icon to hear the response
Click the refresh button (bottom right) to generate a new random word card.
This project is for educational purposes.
Word list data sourced from english-words.
- Mistral AI - Text generation
- DeepSeek - Text generation
- Google Gemini - Text and image generation
- Amazon Polly - Text-to-speech
- Deepgram - Text-to-speech
- dwyl/english-words - Word dictionary
