فان‌فلوئنآموزش

از هوش مصنوعی بخواه سطحت رو بسنجه؛ پرامپت تعیین سطح مکالمه

با AI سطح تقریبی مکالمه‌ات را از چند نمونه بسنج، evidence بگیر، CEFR را فقط برای جهت‌گیری استفاده کن و نتیجه را با retry چک کن.

پاسخ کوتاه

AI می‌تونه از روی چند نمونهٔ مکالمه یک تخمین تمرینی از speaking level بده، اما فقط وقتی ازش evidence، uncertainty و retry بخوای؛ یک label تنها، تعیین سطح رسمی نیست.

سه جمله گفتی. AI می‌گه: «سطحت B2ـه.»

سؤال بعدی نباید این باشه که «جدی؟ 😍»؛ باید بگی: «کجای جواب‌هام evidence این levelه؟»

برای تعیین سطح انگلیسی با هوش مصنوعی، این loop رو نگه دار:

3 SAMPLES → 5 DIMENSIONS → ROUGH RANGE → CAN-DO CHECK → UNCERTAINTY → RETRY

این prompt پایه رو copy کن:

“Assess my English speaking approximately for practice only. Do not treat the result as an official CEFR, IELTS, TOEFL, school, immigration, or hiring score. First collect three short speaking samples on different tasks. Then give me: - a rough speaking level range, not a precise score; - evidence from my answers; - separate notes on fluency, grammar control, vocabulary range, clarity/intelligibility, and interaction; - a cautious comparison with CEFR-style spoken interaction/production can-do descriptions; - what you are uncertain about, especially if the transcript may be wrong; - one easier retry and one harder retry.”

چرا یک جواب برای level دادن کافی نیست؟

چون ممکنه اون یک جواب رو از قبل بلد باشی، topic خیلی آشنا باشه یا transcript تمیزتر از چیزی که واقعاً گفتی دربیاد. یه self-introduction حفظی می‌تونه خیلی قوی به‌نظر بیاد؛ اولین follow-up غیرمنتظره ممکنه تصویر کاملاً متفاوتی بده.

پس قبل از اینکه AI چیزی شبیه A2 یا B1 یا B2 بگه، ازش بخواه حداقل سه sample متفاوت جمع کنه.

سه sample بگیر؛ سه job متفاوت

Original practice example

Tell me what you usually do after work and why.

یک topic آشنا: آیا می‌تونی چند جملهٔ پیوسته بگی و یک reason اضافه کنی؟

Sample 1 برای routine و familiar speaking.

۳۰ تا ۴۵ ثانیه جواب بده؛ script نخون.

FunFluen فارسی.

Original practice example

Tell me about something that went wrong recently and what you did next.

past narration، sequencing و repair رو وارد بازی می‌کنه.

Sample 2 برای تجربهٔ گذشته.

beginning → problem → action → result.

Original practice example

Do you think working from home is better than working in an office? Why?

opinion + reason می‌خواد؛ فقط description ساده نیست.

Sample 3 برای explanation و viewpoint.

یک نظر بده و دو دلیل کوتاه اضافه کن.

Original practice example

What would make you change your mind?

این follow-up interaction رو تست می‌کنه؛ نه فقط monologue آماده.

stress test بعد از Sample 3.

بدون تکرار word-for-word جواب قبلی، ۲۰–۳۰ ثانیه جواب بده.

از AI بخواه level رو به پنج بخش بشکنه

«سطحت B1ـه» اطلاعات کمی می‌ده. این پنج dimension رو جدا بخواه:

پنج بخش برای ارزیابی مکالمه انگلیسی با هوش مصنوعی
DimensionAI دنبال چه evidenceی بگرده؟حواست به چی باشه؟
Fluencyآیا answer رو می‌تونی چند جمله ادامه بدی؟ آیا زیاد restart می‌کنی؟مکث به‌تنهایی level نیست؛ topic difficulty هم مهمه.
Grammar controlآیا tense، agreement و structureها معمولاً meaning رو نگه می‌دارن؟یک slip منفرد رو با pattern تکراری یکی نکن.
Vocabulary rangeآیا برای explanation، opinion و detail word/phrase کافی داری؟rare word داشتن مساوی level بالاتر نیست.
Clarity / intelligibilityآیا message قابل‌فهمه؟اگه AI فقط transcript داره، ازش pronunciation score قطعی نخواه.
Interactionآیا follow-up رو می‌فهمی، clarify می‌کنی و turn رو ادامه می‌دی؟این بخش با monologue تنها دیده نمی‌شه.

CEFR رو برای orientation استفاده کن، نه برای مُهر رسمی

Council of Europe سطح‌های CEFR رو از A1 تا C2 با can-do descriptorها توضیح می‌ده. نکتهٔ مهم اینه که proficiency به skill و activity هم شکسته می‌شه؛ self-assessment grid رسمی CEFR مثلاً spoken interaction و spoken production رو جدا از reading و writing می‌بینه.

پس اگه این صفحه فقط speaking تو رو بررسی کرده، نتیجه رو «speaking estimate» ببین، نه «سطح کل انگلیسی من».

  • A1–A2-ish: phrases و sentenceهای ساده برای موضوع‌های خیلی آشنا و exchangeهای روتین.
  • B1-ish: connected description از experience، plan و opinion با reasonهای ساده.
  • B2-ish: explanation/detail بیشتر، viewpoint روشن‌تر و interaction پایدارتر روی موضوع‌های متنوع.
  • C-level territory: handling پیچیدگی، nuance، flexibility و structure پیشرفته‌تر—که با چند سؤال کوتاه نباید سریع بهش حکم داد.

این paraphrase فقط برای orientationه؛ برای descriptor دقیق، همون صفحهٔ رسمی CEFR رو ببین.

پرامپت کامل تعیین سطح مکالمه با AI

“I want a rough speaking diagnosis, not an official score. First, ask me three speaking tasks: 1) a familiar daily-life question; 2) a past-event question; 3) an opinion question followed by one spontaneous follow-up. Wait for each answer before continuing. After all samples, give me: - a rough speaking range such as ‘around A2–B1’ only if the evidence supports it; - 2–3 pieces of evidence from my answers for fluency; - 2–3 pieces of evidence for grammar control; - evidence for vocabulary range; - a note on clarity/intelligibility, but do not infer pronunciation certainty from transcript alone; - a note on interaction and follow-up handling; - what you are uncertain about; - one easier task and one harder task to challenge the estimate. Do not give IELTS, TOEFL, immigration, school, hiring, or official CEFR certification claims. Do not use decimal scores like B1.7.”

یک sample exchange؛ level باید از evidence بیاد

این مثال writer-createdـه و فقط نشون می‌ده feedback خوب چه شکلیه.

AI: Tell me about a change you made last year.

Learner: Last year I have changed my job because I wanted more free time. At first it was difficult, but now I like my new schedule.

AI: What was the hardest part of changing jobs?

Learner: The hardest part was learning new systems and meeting new people. I needed some weeks to feel comfortable.

Rough speaking estimate: around A2–B1 for these samples, with uncertainty.

Evidence: you can connect several sentences, describe a past experience, give a reason, and answer a follow-up. Grammar control is less stable when talking about completed past events.

Uncertainty: these samples are short and do not show performance on unfamiliar or abstract topics.

نکته: این label illustrativeـه؛ تو باید estimate خودت رو فقط از sampleهای خودت بگیری.

یک correction داخل همین sample

Original: “Last year I have changed my job.”

Classification: برای event تموم‌شده در زمان مشخصِ گذشته، در English استاندارد غلطه.

شنونده چی می‌فهمه؟ می‌فهمه learner سال قبل شغلش رو عوض کرده.

منظور learner: «سال پیش شغلم رو عوض کردم.»

Natural alternative: “Last year I changed my job.”

Context note: “I have changed my job” می‌تونه بدون finished-time marker و وقتی result فعلی مهمه درست باشه؛ اما last year معمولاً simple past می‌خواد.

بعد از مکالمه با هوش مصنوعی خودت رو چطور بسنجی؟ این چک‌لیست رو بزن

اینجا مهم‌ترین بخش assessment شروع می‌شه: خودت باید quality جواب AI رو هم review کنی.

چک‌لیست ارزیابی مکالمه انگلیسی با هوش مصنوعی

Estimate رو challenge کن: یک task آسون‌تر، یک task سخت‌تر

Easier retry

“Describe your usual morning routine for 30 seconds. Use simple sentences and one reason.”

اگه این task خیلی stableـه، baseline پایین‌ترت رو می‌بینی.

Harder retry

“Do you think people should work four days a week? Give your view, two reasons, one disadvantage, and answer one follow-up.”

اینجا explanation، organization و interaction فشار بیشتری می‌گیرن.

“Revise the rough range only if the new evidence changes it. Explain exactly what changed.”

سه session بهتر از یک scoreـه

یک estimate رو دوباره در روز یا topic دیگه امتحان کن. چیزی که تکرار می‌شه ارزش بیشتری از یک slip داره.

لاگ سادهٔ evidence در چند attempt
AttemptTaskPattern تکراریUncertaintyNext target
1daily routineجواب پیوسته، vocabulary محدودکمexpand detail
2past eventpast tense instabilityمتوسطpast narration
3opinion + follow-upreason خوب، follow-up سخت‌ترکمinteraction

اگه با ChatGPT Voice نمونه می‌گیری، transcript رو داور نهایی نکن

طبق مستندات فعلی OpenAI، Voice مکالمهٔ زنده رو پشتیبانی می‌کنه، ولی transcript verbatim نیست و مخصوصاً با overlap یا background noise ممکنه با چیزی که واقعاً گفتی فرق داشته باشه. session هم می‌تونه با usage limit، maximum session length یا context limit تموم بشه.

  • اگه ممکنه recording یا notes خودت رو نگه دار؛
  • transcript mismatch رو pronunciation error قطعی حساب نکن؛
  • از یک transcript تمیز نتیجه نگیر که delivery حتماً عالی بوده؛
  • اگر AI confidently اشتباه کرد، همون confidence رو evidence حساب نکن.

این CEFR certification نیست؛ IELTS/TOEFL score هم نیست

Council of Europe خودش هم تأکید می‌کنه که نقش این نهاد validate کردن کیفیتِ هر ادعای ارتباط بین exam/diploma و CEFR نیست. assessment رسمی طراحی، معیار و validation خودش رو می‌خواد.

«برای sampleهای speaking من، evidence فعلی بیشتر به این range نزدیکه و این دو skill باید بعدی تمرین بشن.»

نه این‌طور:

«من رسماً B2 هستم.»

AI باید دلیل بیاره، نه حکم

اگه آخر session فقط یه badge مثل B1 یا B2 گرفتی، هنوز diagnosis کامل نشده.

سه sample. پنج dimension. یک range. یک uncertainty. یک retry.

level خوب اونیه که بتونی ازش next action دربیاری—نه اینکه فقط توی bio ذهنت بچسبونیش.

منابع