از هوش مصنوعی بخواه سطحت رو بسنجه؛ پرامپت تعیین سطح مکالمه
با AI سطح تقریبی مکالمهات را از چند نمونه بسنج، evidence بگیر، CEFR را فقط برای جهتگیری استفاده کن و نتیجه را با retry چک کن.
AI میتونه از روی چند نمونهٔ مکالمه یک تخمین تمرینی از speaking level بده، اما فقط وقتی ازش evidence، uncertainty و retry بخوای؛ یک label تنها، تعیین سطح رسمی نیست.
سه جمله گفتی. AI میگه: «سطحت B2ـه.»
سؤال بعدی نباید این باشه که «جدی؟ 😍»؛ باید بگی: «کجای جوابهام evidence این levelه؟»
برای تعیین سطح انگلیسی با هوش مصنوعی، این loop رو نگه دار:
3 SAMPLES → 5 DIMENSIONS → ROUGH RANGE → CAN-DO CHECK → UNCERTAINTY → RETRY
این prompt پایه رو copy کن:
“Assess my English speaking approximately for practice only. Do not treat the result as an official CEFR, IELTS, TOEFL, school, immigration, or hiring score. First collect three short speaking samples on different tasks. Then give me: - a rough speaking level range, not a precise score; - evidence from my answers; - separate notes on fluency, grammar control, vocabulary range, clarity/intelligibility, and interaction; - a cautious comparison with CEFR-style spoken interaction/production can-do descriptions; - what you are uncertain about, especially if the transcript may be wrong; - one easier retry and one harder retry.”
چرا یک جواب برای level دادن کافی نیست؟
چون ممکنه اون یک جواب رو از قبل بلد باشی، topic خیلی آشنا باشه یا transcript تمیزتر از چیزی که واقعاً گفتی دربیاد. یه self-introduction حفظی میتونه خیلی قوی بهنظر بیاد؛ اولین follow-up غیرمنتظره ممکنه تصویر کاملاً متفاوتی بده.
پس قبل از اینکه AI چیزی شبیه A2 یا B1 یا B2 بگه، ازش بخواه حداقل سه sample متفاوت جمع کنه.
سه sample بگیر؛ سه job متفاوت
Original practice example
Tell me what you usually do after work and why.
یک topic آشنا: آیا میتونی چند جملهٔ پیوسته بگی و یک reason اضافه کنی؟
Sample 1 برای routine و familiar speaking.
۳۰ تا ۴۵ ثانیه جواب بده؛ script نخون.
Original practice example
Tell me about something that went wrong recently and what you did next.
past narration، sequencing و repair رو وارد بازی میکنه.
Sample 2 برای تجربهٔ گذشته.
beginning → problem → action → result.
Original practice example
Do you think working from home is better than working in an office? Why?
opinion + reason میخواد؛ فقط description ساده نیست.
Sample 3 برای explanation و viewpoint.
یک نظر بده و دو دلیل کوتاه اضافه کن.
Original practice example
What would make you change your mind?
این follow-up interaction رو تست میکنه؛ نه فقط monologue آماده.
stress test بعد از Sample 3.
بدون تکرار word-for-word جواب قبلی، ۲۰–۳۰ ثانیه جواب بده.
از AI بخواه level رو به پنج بخش بشکنه
«سطحت B1ـه» اطلاعات کمی میده. این پنج dimension رو جدا بخواه:
| Dimension | AI دنبال چه evidenceی بگرده؟ | حواست به چی باشه؟ |
|---|---|---|
| Fluency | آیا answer رو میتونی چند جمله ادامه بدی؟ آیا زیاد restart میکنی؟ | مکث بهتنهایی level نیست؛ topic difficulty هم مهمه. |
| Grammar control | آیا tense، agreement و structureها معمولاً meaning رو نگه میدارن؟ | یک slip منفرد رو با pattern تکراری یکی نکن. |
| Vocabulary range | آیا برای explanation، opinion و detail word/phrase کافی داری؟ | rare word داشتن مساوی level بالاتر نیست. |
| Clarity / intelligibility | آیا message قابلفهمه؟ | اگه AI فقط transcript داره، ازش pronunciation score قطعی نخواه. |
| Interaction | آیا follow-up رو میفهمی، clarify میکنی و turn رو ادامه میدی؟ | این بخش با monologue تنها دیده نمیشه. |
CEFR رو برای orientation استفاده کن، نه برای مُهر رسمی
Council of Europe سطحهای CEFR رو از A1 تا C2 با can-do descriptorها توضیح میده. نکتهٔ مهم اینه که proficiency به skill و activity هم شکسته میشه؛ self-assessment grid رسمی CEFR مثلاً spoken interaction و spoken production رو جدا از reading و writing میبینه.
پس اگه این صفحه فقط speaking تو رو بررسی کرده، نتیجه رو «speaking estimate» ببین، نه «سطح کل انگلیسی من».
- A1–A2-ish: phrases و sentenceهای ساده برای موضوعهای خیلی آشنا و exchangeهای روتین.
- B1-ish: connected description از experience، plan و opinion با reasonهای ساده.
- B2-ish: explanation/detail بیشتر، viewpoint روشنتر و interaction پایدارتر روی موضوعهای متنوع.
- C-level territory: handling پیچیدگی، nuance، flexibility و structure پیشرفتهتر—که با چند سؤال کوتاه نباید سریع بهش حکم داد.
این paraphrase فقط برای orientationه؛ برای descriptor دقیق، همون صفحهٔ رسمی CEFR رو ببین.
پرامپت کامل تعیین سطح مکالمه با AI
“I want a rough speaking diagnosis, not an official score. First, ask me three speaking tasks: 1) a familiar daily-life question; 2) a past-event question; 3) an opinion question followed by one spontaneous follow-up. Wait for each answer before continuing. After all samples, give me: - a rough speaking range such as ‘around A2–B1’ only if the evidence supports it; - 2–3 pieces of evidence from my answers for fluency; - 2–3 pieces of evidence for grammar control; - evidence for vocabulary range; - a note on clarity/intelligibility, but do not infer pronunciation certainty from transcript alone; - a note on interaction and follow-up handling; - what you are uncertain about; - one easier task and one harder task to challenge the estimate. Do not give IELTS, TOEFL, immigration, school, hiring, or official CEFR certification claims. Do not use decimal scores like B1.7.”
یک sample exchange؛ level باید از evidence بیاد
این مثال writer-createdـه و فقط نشون میده feedback خوب چه شکلیه.
AI: Tell me about a change you made last year.
Learner: Last year I have changed my job because I wanted more free time. At first it was difficult, but now I like my new schedule.
AI: What was the hardest part of changing jobs?
Learner: The hardest part was learning new systems and meeting new people. I needed some weeks to feel comfortable.
Rough speaking estimate: around A2–B1 for these samples, with uncertainty.
Evidence: you can connect several sentences, describe a past experience, give a reason, and answer a follow-up. Grammar control is less stable when talking about completed past events.
Uncertainty: these samples are short and do not show performance on unfamiliar or abstract topics.
نکته: این label illustrativeـه؛ تو باید estimate خودت رو فقط از sampleهای خودت بگیری.
یک correction داخل همین sample
Original: “Last year I have changed my job.”
Classification: برای event تمومشده در زمان مشخصِ گذشته، در English استاندارد غلطه.
شنونده چی میفهمه؟ میفهمه learner سال قبل شغلش رو عوض کرده.
منظور learner: «سال پیش شغلم رو عوض کردم.»
Natural alternative: “Last year I changed my job.”
Context note: “I have changed my job” میتونه بدون finished-time marker و وقتی result فعلی مهمه درست باشه؛ اما last year معمولاً simple past میخواد.
بعد از مکالمه با هوش مصنوعی خودت رو چطور بسنجی؟ این چکلیست رو بزن
اینجا مهمترین بخش assessment شروع میشه: خودت باید quality جواب AI رو هم review کنی.
چکلیست ارزیابی مکالمه انگلیسی با هوش مصنوعی
Estimate رو challenge کن: یک task آسونتر، یک task سختتر
Easier retry
“Describe your usual morning routine for 30 seconds. Use simple sentences and one reason.”
اگه این task خیلی stableـه، baseline پایینترت رو میبینی.
Harder retry
“Do you think people should work four days a week? Give your view, two reasons, one disadvantage, and answer one follow-up.”
اینجا explanation، organization و interaction فشار بیشتری میگیرن.
“Revise the rough range only if the new evidence changes it. Explain exactly what changed.”
سه session بهتر از یک scoreـه
یک estimate رو دوباره در روز یا topic دیگه امتحان کن. چیزی که تکرار میشه ارزش بیشتری از یک slip داره.
| Attempt | Task | Pattern تکراری | Uncertainty | Next target |
|---|---|---|---|---|
| 1 | daily routine | جواب پیوسته، vocabulary محدود | کم | expand detail |
| 2 | past event | past tense instability | متوسط | past narration |
| 3 | opinion + follow-up | reason خوب، follow-up سختتر | کم | interaction |
اگه با ChatGPT Voice نمونه میگیری، transcript رو داور نهایی نکن
طبق مستندات فعلی OpenAI، Voice مکالمهٔ زنده رو پشتیبانی میکنه، ولی transcript verbatim نیست و مخصوصاً با overlap یا background noise ممکنه با چیزی که واقعاً گفتی فرق داشته باشه. session هم میتونه با usage limit، maximum session length یا context limit تموم بشه.
- اگه ممکنه recording یا notes خودت رو نگه دار؛
- transcript mismatch رو pronunciation error قطعی حساب نکن؛
- از یک transcript تمیز نتیجه نگیر که delivery حتماً عالی بوده؛
- اگر AI confidently اشتباه کرد، همون confidence رو evidence حساب نکن.
این CEFR certification نیست؛ IELTS/TOEFL score هم نیست
Council of Europe خودش هم تأکید میکنه که نقش این نهاد validate کردن کیفیتِ هر ادعای ارتباط بین exam/diploma و CEFR نیست. assessment رسمی طراحی، معیار و validation خودش رو میخواد.
«برای sampleهای speaking من، evidence فعلی بیشتر به این range نزدیکه و این دو skill باید بعدی تمرین بشن.»
نه اینطور:
«من رسماً B2 هستم.»
AI باید دلیل بیاره، نه حکم
اگه آخر session فقط یه badge مثل B1 یا B2 گرفتی، هنوز diagnosis کامل نشده.
سه sample. پنج dimension. یک range. یک uncertainty. یک retry.
level خوب اونیه که بتونی ازش next action دربیاری—نه اینکه فقط توی bio ذهنت بچسبونیش.
منابع
- Council of Europe — The CEFR Levels؛ A1–C2 و can-do descriptorهای skill-specific.
- Council of Europe — CEFR Self-assessment grid؛ spoken interaction و spoken production بهعنوان skillهای جدا.
- Council of Europe — The framework؛ مرز validation/certification و CEFR linkage.
- OpenAI Help Center — ChatGPT Voice؛ Voice conversation، transcript caveat و session/usage limits.