Job Description
for intro paragraphs and general text - Use
for section headers
- Use - for list items
- No CSS classes, inline styles, ,
- Preserve ALL original content word-for-word — do not add, remove, summarise, or rephrase
- If text has no clear sections, just wrap paragraphs in tags
- Return ONLY the HTML. No markdown fences, no explanation, no preamble.
Then there's a "Job description:" section with a lot of text.
So the task is: Take the provided job description text, clean it according to the rules (but the rules here are about HTML formatting, not the earlier "remove social sharing etc." - wait, the very first part describes a different role "job listing editor" task, but then the actual prompt says "You receive a plain-text job description. Your ONLY job is to format it as clean semantic HTML." and gives HTML formatting rules. The initial part about removing social sharing etc. might be context from a system prompt, but the actual instruction is the HTML formatting one. I need to follow the most recent/active instructions.)
Let me re-read the whole user message carefully:
"You are a job listing editor. You receive raw scraped text from a job website.
Your job is to extract ONLY the actual job information and return it in a clean, readable format.
REMOVE all of the following:
- Social sharing text (Share, Facebook, Twitter, WhatsApp, Telegram, etc.)
- Author bylines (by AuthorName, written by, posted by)
- "You may also like" / related posts sections
- Navigation text (Previous, Next, Read more, Back to)
- Comments sections
- Advertisements or promotions
- Cookie/subscription notices
- Duplicate text (if the title appears multiple times, keep it once)
- Category labels that aren't part of the job (e.g., "Internships" as a standalone line)
- Stray numbers or formatting artifacts
KEEP and STRUCTURE:
- Job title
- Company/Organisation name
- Location
- Job description and responsibilities
- Requirements/qualifications
- How to apply / application instructions
- Salary (if mentioned)
- Deadline / closing date
- Job ID / reference number
- Contract type / level
Return ONLY the cleaned job text. Do not add any commentary, headers like "Here is the cleaned version", or markdown formatting. Just the clean job listing text."
Then: "You receive a plain-text job description. Your ONLY job is to format it as clean semantic HTML.
Rules:
- Use
for introductory paragraphs and general text
- Use
for section headers (e.g. Responsibilities, Qualifications, Requirements, Education, Experience, Skills, How to Apply, Key Duties, About the Role)
- Use - for list items (duties, requirements, qualifications, skills)
- No CSS classes, no inline styles, no s, no
tags
- Preserve ALL original content word-for-word — do not add, remove, summarise, or rephrase
- If text has no clear sections, just wrap paragraphs in tags
- Return ONLY the HTML. No markdown fences, no explanation, no preamble.
Job description:
[long text]"
Okay, there are two slightly different instruction sets. The first one is about extracting and cleaning job info, removing fluff, keeping specific fields. The second is about formatting as clean semantic HTML, preserving all content word-for-word, using specific tags.
Which one to follow? The user message starts with "You are a job listing editor..." then gives extraction rules, then says "Return ONLY the cleaned job text." Then it says "You receive a plain-text job description. Your ONLY job is to format it as clean semantic HTML." and gives HTML formatting rules, and then provides the job description text.
I think the actual task I need to perform is the HTML formatting one, because it includes the actual job description text and specific HTML rules. The first part might be a system prompt or context, but the concrete prompt is the HTML formatting one. However, both are in the same message. I need to reconcile.
Let me read the very end: "Job description: [text]" and before that "Rules: ..." The HTML formatting rules say "Preserve ALL original content word-for-word — do not add, remove, summarise, or rephrase". The extraction rules say "REMOVE all of the following: ..." which includes removing social sharing, bylines, etc. The HTML rules don't explicitly say to remove those, but they say preserve all original content word-for-word. There's a conflict.
Actually, the HTML rules say "If text has no clear sections, just wrap paragraphs in
tags" and list specific tags. They don't mention removing anything. The extraction rules explicitly say to remove many things. The job description text included has things like "
", "Employee Contract Type:", "World Vision is a Christian humanitarian organization...", "World Vision Kenya:" disclaimer, etc.
I need to decide: Should I first clean the text per the extraction rules (remove social sharing, bylines, etc.), then format as HTML per the HTML rules? Or should I just format the given text as HTML, preserving everything word-for-word?
The user message structure: It starts with "You are a job listing editor..." (extraction rules), then "You receive a plain-text job description. Your ONLY job is to format it as clean semantic HTML." (HTML rules), then "Job description: [text]". It's possible the user wants me to do both: first clean the text (remove the stuff listed), then format the remaining content as HTML. But the HTML rules say "Preserve ALL original content word-for-word — do not add, remove, summarise, or rephrase". That contradicts removing things.
Maybe the intent is: The extraction rules are the general job of the persona, but for this specific task, I'm to format the given text as HTML, and the "preserve all original content word-for-word" means don't alter the text content, just wrap it. But the extraction rules list things to remove. Which takes precedence?
Let me look at the actual text provided. It has:
- `With 75 years of experience...
` - content
- `` - empty paragraph
- `Local - Fixed Term Employee (Fixed Term)` - contract type
- `Job Description:
` - heading
- Job description paragraphs
- ` ` - italic bold with non-breaking space
- "Please Note:" section about Christian mission, devotions, etc.
- `Job Responsibilities:
` and ul/li lists
- Many more ul/li lists for responsibilities, qualifications, education, World Vision Kenya disclaimer, etc.
- `World Vision Kenya:
` and ul/li about recruitment
- `World Vision Kenya is part of the Inter-Agency Scheme...
`
- `Applicant Types Accepted:
Local Applicants Only`
The HTML rules say use for section headers, - for lists,
for intro/general text. No
, no
, no inline CSS. Preserve all content word-for-word.
If I just wrap everything as HTML per the rules, I need to keep all text exactly as is, just change tags. But there are `
` tags in the original? Actually the text has `
`, `
`, etc. The rules say "No ... no
tags". So I need to handle line breaks without
. Probably wrap in
or just let paragraphs flow.
Also, the extraction rules say remove things like social sharing, bylines, etc. The text doesn't seem to have those explicitly, but has the "World Vision Kenya:" disclaimer and "Please Note:" which might be considered part of the job or not. The extraction rules also say "Remove duplicate text (if the title appears multiple times, keep it once)" - not sure about title duplication.
Given the conflict, I need to interpret the prompt as a whole. The user says: "You receive a plain-text job description. Your ONLY job is to format it as clean semantic HTML." Then gives rules. Then gives the job description. The HTML rules are the specific instructions for this task. The extraction rules might be background/context from the system, but the actual command is the HTML formatting one. However, the extraction rules explicitly say "REMOVE all of the following:" and list things. If I follow the extraction rules, I'd remove parts of the job description that might be considered essential (like the "Please Note:" Christian mission statement, the disclaimer, etc.). But the HTML rules say "Preserve ALL original content word-for-word".
Maybe the correct approach is: First, apply the extraction rules to remove the listed fluff, then format the remaining clean text as HTML per the HTML rules. But the HTML rules say "Preserve ALL original content word-for-word — do not add, remove, summarise, or rephrase". If I remove things first, I'm not preserving all original content of the given text, but I'm following the extraction rules which are part of the persona's job.
Let me re-read the very beginning: "You are a job listing editor. You receive raw scraped text from a job website. Your job is to extract ONLY the actual job information and return it in a clean, readable format." Then extraction rules. Then "Return ONLY the cleaned job text. Do not add any commentary, headers like "Here is the cleaned version", or markdown formatting. Just the clean job listing text."
Then: "You receive a plain-text job description. Your ONLY job is to format it as clean semantic HTML." This seems like a separate instruction or a reformulation. Maybe the user is giving me two prompts in one, or the second is the actual task format.
Given the ambiguity, I should look at the job description text and see what makes sense. The text is a job listing for an M&E Officer at World Vision Kenya. It has a lot of content. The extraction rules would remove things like the "Share" etc., but those aren't present. It has the "Please Note:" about Christian mission, which might be considered part of the job or not. The "World Vision Kenya:" disclaimer is standard but might be removed per "category labels that aren't part of the job". The "Applicant Types Accepted: Local Applicants Only" might be kept or removed.
The HTML formatting rules: "Use
for introductory paragraphs and general text", "Use
for section headers", "Use - for list items", "No CSS classes, no inline styles, no s, no
tags", "Preserve ALL original content word-for-word — do not add, remove, summarise, or rephrase", "If text has no clear sections, just wrap paragraphs in tags", "Return ONLY the HTML. No markdown fences, no explanation, no preamble."
I think the safest interpretation is: The user wants me to take the given job description text and format it as clean semantic HTML, following the tag rules, preserving all content word-for-word. The extraction rules might be the general persona description, but the specific task here is the HTML formatting, and the "preserve all original content word-for-word" overrides or specifies that I shouldn't remove content unless it's clearly fluff like social sharing etc. But the text doesn't have explicit social sharing, author bylines, etc. It has structural elements like `
`, `
`, etc.
Actually, re-reading carefully: The user message has two parts. The first part is "You are a job listing editor..." with extraction rules, and says "Return ONLY the cleaned job text." The second part starts with "You receive a plain-text job description. Your ONLY job is to format it as clean semantic HTML." and has HTML rules, and then "Job description: [text]". It's possible the user copied two different prompts, but I need to output something.
Maybe the intended task is: Apply the extraction rules to clean the text, then format the cleaned text as HTML. But the HTML rules say "Preserve ALL original content word-for-word", which contradicts removing things. Unless "original content" refers to the content after extraction.
Let me think differently: Perhaps the extraction rules are the primary task, and the HTML formatting is just a way to present the cleaned text. But the extraction rules say "Return ONLY the cleaned job text." not HTML. Then the second part says format as HTML. Maybe I need to do both: clean it (remove the listed items), then output the cleaned text as HTML with the specified tags, preserving the cleaned content word-for-word.
Given the job description text provided, if I remove things per extraction rules:
- Remove social sharing: not present.
- Remove author bylines: not present (no "by AuthorName").
- Remove "You may also like": not present.
- Remove navigation: not present.
- Remove comments: not present.
- Remove ads/promotions: not present.
- Remove cookie/subscription notices: not present.
- Remove duplicate text: maybe the title? Not sure.
- Remove category labels not part of job: maybe "World Vision Kenya:" disclaimer section? But it's part of the job listing.
- Remove stray numbers/formatting artifacts: the text has some.
If I just format as HTML per the HTML rules, I need to handle the existing tags. The text has ``, ``, ``, ``, `
`, ``, `- `. The HTML rules say "Use
for introductory paragraphs and general text", "Use
for section headers", "Use - for list items", "No ... no
tags". So I need to convert the text to use only those tags, removing
, , , ? No, "Preserve ALL original content word-for-word" means keep the text content exactly, but I can change the tags. `` and `` and `` are formatting markup, but the rule says "Preserve ALL original content word-for-word — do not add, remove, summarise, or rephrase". Does that mean keep the literal text including the tags, or keep the visible text and re-tag it? Usually in such tasks, "content" means the visible text, and you re-markup it. But it says "do not add, remove, summarise, or rephrase", which suggests don't change the words themselves.
This is tricky. Let me look at the exact wording: "Preserve ALL original content word-for-word — do not add, remove, summarise, or rephrase". If I wrap paragraphs in , I'm not removing or adding content, just tagging. If I keep , , tags, those are not "content" per se, but markup. The rule says "No CSS classes, no inline styles, no s, no
tags". It explicitly bans
tags. So I must remove
tags and presumably replace them with or just let paragraphs flow. It also says no
s. It doesn't explicitly ban , , , but says "Use for section headers", "Use - for list items". It might imply that I should use those tags for structure, and the text content goes inside them.
Given the complexity, maybe the best approach is to output the HTML version of the job description, following the HTML rules as closely as possible, while preserving all visible text word-for-word, and removing/handling the banned tags (
, , inline CSS, etc.). I'll keep , , maybe, or I can just keep the text and use etc. if needed, but the rules say "Use for section headers (e.g. Responsibilities, Qualifications, ...)". The text has `
Job Responsibilities:
` etc. I could convert those to Job Responsibilities
but the rule says "Preserve ALL original content word-for-word". If I change "Job Responsibilities:" to an , I'm not removing content, but I'm changing the tag. The content "Job Responsibilities:" would still be there.
Actually, maybe I should just wrap everything in
and
- as appropriate, and remove
by merging lines into paragraphs. The text has `
`, `
`. I'll treat each `
...
` block as a paragraph, and each `- ...
` as a list. I'll keep the text inside exactly as is. I'll remove `
` tags by not including them, and just let paragraphs be separate. I'll keep
Apply Now ↗
How well do you match?
Get an instant AI match score for this role — free, takes 3 minutes.
Tailor your CV for this role
The concierge rewrites your whole CV and writes a matching cover letter for this job — opens right here, nothing to paste.
Tailor My CV to This Job ✍️
Free cover letter for this job
Upload your CV and get a tailored cover letter in seconds — free, no account needed.
Generate a Cover Letter 📝
My Job Concierge
Let jobs find you
Leave your email and our AI matches you to new jobs across 24 African markets — free. You wait for the call.
Almost there
Add your CV for real matches
Upload your CV and we score every new job against your real experience — only strong matches reach your inbox. Optional, but it makes your matches far sharper.
You're in.
We'll email you the moment a job matches your profile. Check your inbox for a welcome from My Job Concierge.
ECHO
Your MJC Assistant
I'm ECHO, your MJC career assistant. I can help you find jobs, explore career tools, and connect with opportunities across Africa.
- Preserve ALL original content word-for-word — do not add, remove, summarise, or rephrase - If text has no clear sections, just wrap paragraphs in
tags - Return ONLY the HTML. No markdown fences, no explanation, no preamble. Then there's a "Job description:" section with a lot of text. So the task is: Take the provided job description text, clean it according to the rules (but the rules here are about HTML formatting, not the earlier "remove social sharing etc." - wait, the very first part describes a different role "job listing editor" task, but then the actual prompt says "You receive a plain-text job description. Your ONLY job is to format it as clean semantic HTML." and gives HTML formatting rules. The initial part about removing social sharing etc. might be context from a system prompt, but the actual instruction is the HTML formatting one. I need to follow the most recent/active instructions.) Let me re-read the whole user message carefully: "You are a job listing editor. You receive raw scraped text from a job website. Your job is to extract ONLY the actual job information and return it in a clean, readable format. REMOVE all of the following: - Social sharing text (Share, Facebook, Twitter, WhatsApp, Telegram, etc.) - Author bylines (by AuthorName, written by, posted by) - "You may also like" / related posts sections - Navigation text (Previous, Next, Read more, Back to) - Comments sections - Advertisements or promotions - Cookie/subscription notices - Duplicate text (if the title appears multiple times, keep it once) - Category labels that aren't part of the job (e.g., "Internships" as a standalone line) - Stray numbers or formatting artifacts KEEP and STRUCTURE: - Job title - Company/Organisation name - Location - Job description and responsibilities - Requirements/qualifications - How to apply / application instructions - Salary (if mentioned) - Deadline / closing date - Job ID / reference number - Contract type / level Return ONLY the cleaned job text. Do not add any commentary, headers like "Here is the cleaned version", or markdown formatting. Just the clean job listing text." Then: "You receive a plain-text job description. Your ONLY job is to format it as clean semantic HTML. Rules: - Use
for introductory paragraphs and general text - Use
for section headers (e.g. Responsibilities, Qualifications, Requirements, Education, Experience, Skills, How to Apply, Key Duties, About the Role)
- Use - for list items (duties, requirements, qualifications, skills)
- No CSS classes, no inline styles, no s, no
tags
- Preserve ALL original content word-for-word — do not add, remove, summarise, or rephrase
- If text has no clear sections, just wrap paragraphs in tags
- Return ONLY the HTML. No markdown fences, no explanation, no preamble.
Job description:
[long text]"
Okay, there are two slightly different instruction sets. The first one is about extracting and cleaning job info, removing fluff, keeping specific fields. The second is about formatting as clean semantic HTML, preserving all content word-for-word, using specific tags.
Which one to follow? The user message starts with "You are a job listing editor..." then gives extraction rules, then says "Return ONLY the cleaned job text." Then it says "You receive a plain-text job description. Your ONLY job is to format it as clean semantic HTML." and gives HTML formatting rules, and then provides the job description text.
I think the actual task I need to perform is the HTML formatting one, because it includes the actual job description text and specific HTML rules. The first part might be a system prompt or context, but the concrete prompt is the HTML formatting one. However, both are in the same message. I need to reconcile.
Let me read the very end: "Job description: [text]" and before that "Rules: ..." The HTML formatting rules say "Preserve ALL original content word-for-word — do not add, remove, summarise, or rephrase". The extraction rules say "REMOVE all of the following: ..." which includes removing social sharing, bylines, etc. The HTML rules don't explicitly say to remove those, but they say preserve all original content word-for-word. There's a conflict.
Actually, the HTML rules say "If text has no clear sections, just wrap paragraphs in
tags" and list specific tags. They don't mention removing anything. The extraction rules explicitly say to remove many things. The job description text included has things like "
", "Employee Contract Type:", "World Vision is a Christian humanitarian organization...", "World Vision Kenya:" disclaimer, etc.
I need to decide: Should I first clean the text per the extraction rules (remove social sharing, bylines, etc.), then format as HTML per the HTML rules? Or should I just format the given text as HTML, preserving everything word-for-word?
The user message structure: It starts with "You are a job listing editor..." (extraction rules), then "You receive a plain-text job description. Your ONLY job is to format it as clean semantic HTML." (HTML rules), then "Job description: [text]". It's possible the user wants me to do both: first clean the text (remove the stuff listed), then format the remaining content as HTML. But the HTML rules say "Preserve ALL original content word-for-word — do not add, remove, summarise, or rephrase". That contradicts removing things.
Maybe the intent is: The extraction rules are the general job of the persona, but for this specific task, I'm to format the given text as HTML, and the "preserve all original content word-for-word" means don't alter the text content, just wrap it. But the extraction rules list things to remove. Which takes precedence?
Let me look at the actual text provided. It has:
- `With 75 years of experience...
` - content
- `` - empty paragraph
- `Local - Fixed Term Employee (Fixed Term)` - contract type
- `Job Description:
` - heading
- Job description paragraphs
- ` ` - italic bold with non-breaking space
- "Please Note:" section about Christian mission, devotions, etc.
- `Job Responsibilities:
` and ul/li lists
- Many more ul/li lists for responsibilities, qualifications, education, World Vision Kenya disclaimer, etc.
- `World Vision Kenya:
` and ul/li about recruitment
- `World Vision Kenya is part of the Inter-Agency Scheme...
`
- `Applicant Types Accepted:
Local Applicants Only`
The HTML rules say use for section headers, - for lists,
for intro/general text. No
, no
, no inline CSS. Preserve all content word-for-word.
If I just wrap everything as HTML per the rules, I need to keep all text exactly as is, just change tags. But there are `
` tags in the original? Actually the text has `
`, `
`, etc. The rules say "No ... no
tags". So I need to handle line breaks without
. Probably wrap in
or just let paragraphs flow.
Also, the extraction rules say remove things like social sharing, bylines, etc. The text doesn't seem to have those explicitly, but has the "World Vision Kenya:" disclaimer and "Please Note:" which might be considered part of the job or not. The extraction rules also say "Remove duplicate text (if the title appears multiple times, keep it once)" - not sure about title duplication.
Given the conflict, I need to interpret the prompt as a whole. The user says: "You receive a plain-text job description. Your ONLY job is to format it as clean semantic HTML." Then gives rules. Then gives the job description. The HTML rules are the specific instructions for this task. The extraction rules might be background/context from the system, but the actual command is the HTML formatting one. However, the extraction rules explicitly say "REMOVE all of the following:" and list things. If I follow the extraction rules, I'd remove parts of the job description that might be considered essential (like the "Please Note:" Christian mission statement, the disclaimer, etc.). But the HTML rules say "Preserve ALL original content word-for-word".
Maybe the correct approach is: First, apply the extraction rules to remove the listed fluff, then format the remaining clean text as HTML per the HTML rules. But the HTML rules say "Preserve ALL original content word-for-word — do not add, remove, summarise, or rephrase". If I remove things first, I'm not preserving all original content of the given text, but I'm following the extraction rules which are part of the persona's job.
Let me re-read the very beginning: "You are a job listing editor. You receive raw scraped text from a job website. Your job is to extract ONLY the actual job information and return it in a clean, readable format." Then extraction rules. Then "Return ONLY the cleaned job text. Do not add any commentary, headers like "Here is the cleaned version", or markdown formatting. Just the clean job listing text."
Then: "You receive a plain-text job description. Your ONLY job is to format it as clean semantic HTML." This seems like a separate instruction or a reformulation. Maybe the user is giving me two prompts in one, or the second is the actual task format.
Given the ambiguity, I should look at the job description text and see what makes sense. The text is a job listing for an M&E Officer at World Vision Kenya. It has a lot of content. The extraction rules would remove things like the "Share" etc., but those aren't present. It has the "Please Note:" about Christian mission, which might be considered part of the job or not. The "World Vision Kenya:" disclaimer is standard but might be removed per "category labels that aren't part of the job". The "Applicant Types Accepted: Local Applicants Only" might be kept or removed.
The HTML formatting rules: "Use
for introductory paragraphs and general text", "Use
for section headers", "Use - for list items", "No CSS classes, no inline styles, no s, no
tags", "Preserve ALL original content word-for-word — do not add, remove, summarise, or rephrase", "If text has no clear sections, just wrap paragraphs in tags", "Return ONLY the HTML. No markdown fences, no explanation, no preamble."
I think the safest interpretation is: The user wants me to take the given job description text and format it as clean semantic HTML, following the tag rules, preserving all content word-for-word. The extraction rules might be the general persona description, but the specific task here is the HTML formatting, and the "preserve all original content word-for-word" overrides or specifies that I shouldn't remove content unless it's clearly fluff like social sharing etc. But the text doesn't have explicit social sharing, author bylines, etc. It has structural elements like `
`, `
`, etc.
Actually, re-reading carefully: The user message has two parts. The first part is "You are a job listing editor..." with extraction rules, and says "Return ONLY the cleaned job text." The second part starts with "You receive a plain-text job description. Your ONLY job is to format it as clean semantic HTML." and has HTML rules, and then "Job description: [text]". It's possible the user copied two different prompts, but I need to output something.
Maybe the intended task is: Apply the extraction rules to clean the text, then format the cleaned text as HTML. But the HTML rules say "Preserve ALL original content word-for-word", which contradicts removing things. Unless "original content" refers to the content after extraction.
Let me think differently: Perhaps the extraction rules are the primary task, and the HTML formatting is just a way to present the cleaned text. But the extraction rules say "Return ONLY the cleaned job text." not HTML. Then the second part says format as HTML. Maybe I need to do both: clean it (remove the listed items), then output the cleaned text as HTML with the specified tags, preserving the cleaned content word-for-word.
Given the job description text provided, if I remove things per extraction rules:
- Remove social sharing: not present.
- Remove author bylines: not present (no "by AuthorName").
- Remove "You may also like": not present.
- Remove navigation: not present.
- Remove comments: not present.
- Remove ads/promotions: not present.
- Remove cookie/subscription notices: not present.
- Remove duplicate text: maybe the title? Not sure.
- Remove category labels not part of job: maybe "World Vision Kenya:" disclaimer section? But it's part of the job listing.
- Remove stray numbers/formatting artifacts: the text has some.
If I just format as HTML per the HTML rules, I need to handle the existing tags. The text has ``, ``, ``, ``, `
`, ``, `- `. The HTML rules say "Use
for introductory paragraphs and general text", "Use
for section headers", "Use - for list items", "No ... no
tags". So I need to convert the text to use only those tags, removing
, , , ? No, "Preserve ALL original content word-for-word" means keep the text content exactly, but I can change the tags. `` and `` and `` are formatting markup, but the rule says "Preserve ALL original content word-for-word — do not add, remove, summarise, or rephrase". Does that mean keep the literal text including the tags, or keep the visible text and re-tag it? Usually in such tasks, "content" means the visible text, and you re-markup it. But it says "do not add, remove, summarise, or rephrase", which suggests don't change the words themselves.
This is tricky. Let me look at the exact wording: "Preserve ALL original content word-for-word — do not add, remove, summarise, or rephrase". If I wrap paragraphs in , I'm not removing or adding content, just tagging. If I keep , , tags, those are not "content" per se, but markup. The rule says "No CSS classes, no inline styles, no s, no
tags". It explicitly bans
tags. So I must remove
tags and presumably replace them with or just let paragraphs flow. It also says no
s. It doesn't explicitly ban , , , but says "Use for section headers", "Use - for list items". It might imply that I should use those tags for structure, and the text content goes inside them.
Given the complexity, maybe the best approach is to output the HTML version of the job description, following the HTML rules as closely as possible, while preserving all visible text word-for-word, and removing/handling the banned tags (
, , inline CSS, etc.). I'll keep , , maybe, or I can just keep the text and use etc. if needed, but the rules say "Use for section headers (e.g. Responsibilities, Qualifications, ...)". The text has `
Job Responsibilities:
` etc. I could convert those to Job Responsibilities
but the rule says "Preserve ALL original content word-for-word". If I change "Job Responsibilities:" to an , I'm not removing content, but I'm changing the tag. The content "Job Responsibilities:" would still be there.
Actually, maybe I should just wrap everything in
and
- as appropriate, and remove
by merging lines into paragraphs. The text has `
`, `
`. I'll treat each `
...
` block as a paragraph, and each `- ...
` as a list. I'll keep the text inside exactly as is. I'll remove `
` tags by not including them, and just let paragraphs be separate. I'll keep
Apply Now ↗
How well do you match?
Get an instant AI match score for this role — free, takes 3 minutes.
Tailor your CV for this role
The concierge rewrites your whole CV and writes a matching cover letter for this job — opens right here, nothing to paste.
Tailor My CV to This Job ✍️
Free cover letter for this job
Upload your CV and get a tailored cover letter in seconds — free, no account needed.
Generate a Cover Letter 📝
My Job Concierge
Let jobs find you
Leave your email and our AI matches you to new jobs across 24 African markets — free. You wait for the call.
Almost there
Add your CV for real matches
Upload your CV and we score every new job against your real experience — only strong matches reach your inbox. Optional, but it makes your matches far sharper.
You're in.
We'll email you the moment a job matches your profile. Check your inbox for a welcome from My Job Concierge.
ECHO
Your MJC Assistant
I'm ECHO, your MJC career assistant. I can help you find jobs, explore career tools, and connect with opportunities across Africa.
tags - Preserve ALL original content word-for-word — do not add, remove, summarise, or rephrase - If text has no clear sections, just wrap paragraphs in
tags - Return ONLY the HTML. No markdown fences, no explanation, no preamble. Job description: [long text]" Okay, there are two slightly different instruction sets. The first one is about extracting and cleaning job info, removing fluff, keeping specific fields. The second is about formatting as clean semantic HTML, preserving all content word-for-word, using specific tags. Which one to follow? The user message starts with "You are a job listing editor..." then gives extraction rules, then says "Return ONLY the cleaned job text." Then it says "You receive a plain-text job description. Your ONLY job is to format it as clean semantic HTML." and gives HTML formatting rules, and then provides the job description text. I think the actual task I need to perform is the HTML formatting one, because it includes the actual job description text and specific HTML rules. The first part might be a system prompt or context, but the concrete prompt is the HTML formatting one. However, both are in the same message. I need to reconcile. Let me read the very end: "Job description: [text]" and before that "Rules: ..." The HTML formatting rules say "Preserve ALL original content word-for-word — do not add, remove, summarise, or rephrase". The extraction rules say "REMOVE all of the following: ..." which includes removing social sharing, bylines, etc. The HTML rules don't explicitly say to remove those, but they say preserve all original content word-for-word. There's a conflict. Actually, the HTML rules say "If text has no clear sections, just wrap paragraphs in
tags" and list specific tags. They don't mention removing anything. The extraction rules explicitly say to remove many things. The job description text included has things like "
", "Employee Contract Type:", "World Vision is a Christian humanitarian organization...", "World Vision Kenya:" disclaimer, etc. I need to decide: Should I first clean the text per the extraction rules (remove social sharing, bylines, etc.), then format as HTML per the HTML rules? Or should I just format the given text as HTML, preserving everything word-for-word? The user message structure: It starts with "You are a job listing editor..." (extraction rules), then "You receive a plain-text job description. Your ONLY job is to format it as clean semantic HTML." (HTML rules), then "Job description: [text]". It's possible the user wants me to do both: first clean the text (remove the stuff listed), then format the remaining content as HTML. But the HTML rules say "Preserve ALL original content word-for-word — do not add, remove, summarise, or rephrase". That contradicts removing things. Maybe the intent is: The extraction rules are the general job of the persona, but for this specific task, I'm to format the given text as HTML, and the "preserve all original content word-for-word" means don't alter the text content, just wrap it. But the extraction rules list things to remove. Which takes precedence? Let me look at the actual text provided. It has: - `With 75 years of experience...
` - content - `` - empty paragraph - `Local - Fixed Term Employee (Fixed Term)` - contract type - `Job Description:
` - heading - Job description paragraphs - ` ` - italic bold with non-breaking space - "Please Note:" section about Christian mission, devotions, etc. - `Job Responsibilities:
` and ul/li lists - Many more ul/li lists for responsibilities, qualifications, education, World Vision Kenya disclaimer, etc. - `World Vision Kenya:
` and ul/li about recruitment - `World Vision Kenya is part of the Inter-Agency Scheme...
` - `Applicant Types Accepted:
Local Applicants Only` The HTML rules say use for section headers, - for lists,
for intro/general text. No
, no
, no inline CSS. Preserve all content word-for-word.
If I just wrap everything as HTML per the rules, I need to keep all text exactly as is, just change tags. But there are `
` tags in the original? Actually the text has `
`, `
`, etc. The rules say "No ... no
tags". So I need to handle line breaks without
. Probably wrap in
or just let paragraphs flow.
Also, the extraction rules say remove things like social sharing, bylines, etc. The text doesn't seem to have those explicitly, but has the "World Vision Kenya:" disclaimer and "Please Note:" which might be considered part of the job or not. The extraction rules also say "Remove duplicate text (if the title appears multiple times, keep it once)" - not sure about title duplication.
Given the conflict, I need to interpret the prompt as a whole. The user says: "You receive a plain-text job description. Your ONLY job is to format it as clean semantic HTML." Then gives rules. Then gives the job description. The HTML rules are the specific instructions for this task. The extraction rules might be background/context from the system, but the actual command is the HTML formatting one. However, the extraction rules explicitly say "REMOVE all of the following:" and list things. If I follow the extraction rules, I'd remove parts of the job description that might be considered essential (like the "Please Note:" Christian mission statement, the disclaimer, etc.). But the HTML rules say "Preserve ALL original content word-for-word".
Maybe the correct approach is: First, apply the extraction rules to remove the listed fluff, then format the remaining clean text as HTML per the HTML rules. But the HTML rules say "Preserve ALL original content word-for-word — do not add, remove, summarise, or rephrase". If I remove things first, I'm not preserving all original content of the given text, but I'm following the extraction rules which are part of the persona's job.
Let me re-read the very beginning: "You are a job listing editor. You receive raw scraped text from a job website. Your job is to extract ONLY the actual job information and return it in a clean, readable format." Then extraction rules. Then "Return ONLY the cleaned job text. Do not add any commentary, headers like "Here is the cleaned version", or markdown formatting. Just the clean job listing text."
Then: "You receive a plain-text job description. Your ONLY job is to format it as clean semantic HTML." This seems like a separate instruction or a reformulation. Maybe the user is giving me two prompts in one, or the second is the actual task format.
Given the ambiguity, I should look at the job description text and see what makes sense. The text is a job listing for an M&E Officer at World Vision Kenya. It has a lot of content. The extraction rules would remove things like the "Share" etc., but those aren't present. It has the "Please Note:" about Christian mission, which might be considered part of the job or not. The "World Vision Kenya:" disclaimer is standard but might be removed per "category labels that aren't part of the job". The "Applicant Types Accepted: Local Applicants Only" might be kept or removed.
The HTML formatting rules: "Use
for introductory paragraphs and general text", "Use
for section headers", "Use - for list items", "No CSS classes, no inline styles, no s, no
tags", "Preserve ALL original content word-for-word — do not add, remove, summarise, or rephrase", "If text has no clear sections, just wrap paragraphs in tags", "Return ONLY the HTML. No markdown fences, no explanation, no preamble."
I think the safest interpretation is: The user wants me to take the given job description text and format it as clean semantic HTML, following the tag rules, preserving all content word-for-word. The extraction rules might be the general persona description, but the specific task here is the HTML formatting, and the "preserve all original content word-for-word" overrides or specifies that I shouldn't remove content unless it's clearly fluff like social sharing etc. But the text doesn't have explicit social sharing, author bylines, etc. It has structural elements like `
`, `
`, etc.
Actually, re-reading carefully: The user message has two parts. The first part is "You are a job listing editor..." with extraction rules, and says "Return ONLY the cleaned job text." The second part starts with "You receive a plain-text job description. Your ONLY job is to format it as clean semantic HTML." and has HTML rules, and then "Job description: [text]". It's possible the user copied two different prompts, but I need to output something.
Maybe the intended task is: Apply the extraction rules to clean the text, then format the cleaned text as HTML. But the HTML rules say "Preserve ALL original content word-for-word", which contradicts removing things. Unless "original content" refers to the content after extraction.
Let me think differently: Perhaps the extraction rules are the primary task, and the HTML formatting is just a way to present the cleaned text. But the extraction rules say "Return ONLY the cleaned job text." not HTML. Then the second part says format as HTML. Maybe I need to do both: clean it (remove the listed items), then output the cleaned text as HTML with the specified tags, preserving the cleaned content word-for-word.
Given the job description text provided, if I remove things per extraction rules:
- Remove social sharing: not present.
- Remove author bylines: not present (no "by AuthorName").
- Remove "You may also like": not present.
- Remove navigation: not present.
- Remove comments: not present.
- Remove ads/promotions: not present.
- Remove cookie/subscription notices: not present.
- Remove duplicate text: maybe the title? Not sure.
- Remove category labels not part of job: maybe "World Vision Kenya:" disclaimer section? But it's part of the job listing.
- Remove stray numbers/formatting artifacts: the text has some.
If I just format as HTML per the HTML rules, I need to handle the existing tags. The text has ``, ``, ``, ``, `
`, ``, `- `. The HTML rules say "Use
for introductory paragraphs and general text", "Use
for section headers", "Use - for list items", "No ... no
tags". So I need to convert the text to use only those tags, removing
, , , ? No, "Preserve ALL original content word-for-word" means keep the text content exactly, but I can change the tags. `` and `` and `` are formatting markup, but the rule says "Preserve ALL original content word-for-word — do not add, remove, summarise, or rephrase". Does that mean keep the literal text including the tags, or keep the visible text and re-tag it? Usually in such tasks, "content" means the visible text, and you re-markup it. But it says "do not add, remove, summarise, or rephrase", which suggests don't change the words themselves.
This is tricky. Let me look at the exact wording: "Preserve ALL original content word-for-word — do not add, remove, summarise, or rephrase". If I wrap paragraphs in , I'm not removing or adding content, just tagging. If I keep , , tags, those are not "content" per se, but markup. The rule says "No CSS classes, no inline styles, no s, no
tags". It explicitly bans
tags. So I must remove
tags and presumably replace them with or just let paragraphs flow. It also says no
s. It doesn't explicitly ban , , , but says "Use for section headers", "Use - for list items". It might imply that I should use those tags for structure, and the text content goes inside them.
Given the complexity, maybe the best approach is to output the HTML version of the job description, following the HTML rules as closely as possible, while preserving all visible text word-for-word, and removing/handling the banned tags (
, , inline CSS, etc.). I'll keep , , maybe, or I can just keep the text and use etc. if needed, but the rules say "Use for section headers (e.g. Responsibilities, Qualifications, ...)". The text has `
Job Responsibilities:
` etc. I could convert those to Job Responsibilities
but the rule says "Preserve ALL original content word-for-word". If I change "Job Responsibilities:" to an , I'm not removing content, but I'm changing the tag. The content "Job Responsibilities:" would still be there.
Actually, maybe I should just wrap everything in
and
- as appropriate, and remove
by merging lines into paragraphs. The text has `
`, `
`. I'll treat each `
...
` block as a paragraph, and each `- ...
` as a list. I'll keep the text inside exactly as is. I'll remove `
` tags by not including them, and just let paragraphs be separate. I'll keep
Apply Now ↗
How well do you match?
Get an instant AI match score for this role — free, takes 3 minutes.
Tailor your CV for this role
The concierge rewrites your whole CV and writes a matching cover letter for this job — opens right here, nothing to paste.
Tailor My CV to This Job ✍️
Free cover letter for this job
Upload your CV and get a tailored cover letter in seconds — free, no account needed.
Generate a Cover Letter 📝
for intro/general text. No
, no
` tags in the original? Actually the text has `
`, `
`, etc. The rules say "No ... no
tags". So I need to handle line breaks without
. Probably wrap in
or just let paragraphs flow. Also, the extraction rules say remove things like social sharing, bylines, etc. The text doesn't seem to have those explicitly, but has the "World Vision Kenya:" disclaimer and "Please Note:" which might be considered part of the job or not. The extraction rules also say "Remove duplicate text (if the title appears multiple times, keep it once)" - not sure about title duplication. Given the conflict, I need to interpret the prompt as a whole. The user says: "You receive a plain-text job description. Your ONLY job is to format it as clean semantic HTML." Then gives rules. Then gives the job description. The HTML rules are the specific instructions for this task. The extraction rules might be background/context from the system, but the actual command is the HTML formatting one. However, the extraction rules explicitly say "REMOVE all of the following:" and list things. If I follow the extraction rules, I'd remove parts of the job description that might be considered essential (like the "Please Note:" Christian mission statement, the disclaimer, etc.). But the HTML rules say "Preserve ALL original content word-for-word". Maybe the correct approach is: First, apply the extraction rules to remove the listed fluff, then format the remaining clean text as HTML per the HTML rules. But the HTML rules say "Preserve ALL original content word-for-word — do not add, remove, summarise, or rephrase". If I remove things first, I'm not preserving all original content of the given text, but I'm following the extraction rules which are part of the persona's job. Let me re-read the very beginning: "You are a job listing editor. You receive raw scraped text from a job website. Your job is to extract ONLY the actual job information and return it in a clean, readable format." Then extraction rules. Then "Return ONLY the cleaned job text. Do not add any commentary, headers like "Here is the cleaned version", or markdown formatting. Just the clean job listing text." Then: "You receive a plain-text job description. Your ONLY job is to format it as clean semantic HTML." This seems like a separate instruction or a reformulation. Maybe the user is giving me two prompts in one, or the second is the actual task format. Given the ambiguity, I should look at the job description text and see what makes sense. The text is a job listing for an M&E Officer at World Vision Kenya. It has a lot of content. The extraction rules would remove things like the "Share" etc., but those aren't present. It has the "Please Note:" about Christian mission, which might be considered part of the job or not. The "World Vision Kenya:" disclaimer is standard but might be removed per "category labels that aren't part of the job". The "Applicant Types Accepted: Local Applicants Only" might be kept or removed. The HTML formatting rules: "Use
for introductory paragraphs and general text", "Use
for section headers", "Use - for list items", "No CSS classes, no inline styles, no s, no
tags", "Preserve ALL original content word-for-word — do not add, remove, summarise, or rephrase", "If text has no clear sections, just wrap paragraphs in tags", "Return ONLY the HTML. No markdown fences, no explanation, no preamble."
I think the safest interpretation is: The user wants me to take the given job description text and format it as clean semantic HTML, following the tag rules, preserving all content word-for-word. The extraction rules might be the general persona description, but the specific task here is the HTML formatting, and the "preserve all original content word-for-word" overrides or specifies that I shouldn't remove content unless it's clearly fluff like social sharing etc. But the text doesn't have explicit social sharing, author bylines, etc. It has structural elements like `
`, `
`, etc.
Actually, re-reading carefully: The user message has two parts. The first part is "You are a job listing editor..." with extraction rules, and says "Return ONLY the cleaned job text." The second part starts with "You receive a plain-text job description. Your ONLY job is to format it as clean semantic HTML." and has HTML rules, and then "Job description: [text]". It's possible the user copied two different prompts, but I need to output something.
Maybe the intended task is: Apply the extraction rules to clean the text, then format the cleaned text as HTML. But the HTML rules say "Preserve ALL original content word-for-word", which contradicts removing things. Unless "original content" refers to the content after extraction.
Let me think differently: Perhaps the extraction rules are the primary task, and the HTML formatting is just a way to present the cleaned text. But the extraction rules say "Return ONLY the cleaned job text." not HTML. Then the second part says format as HTML. Maybe I need to do both: clean it (remove the listed items), then output the cleaned text as HTML with the specified tags, preserving the cleaned content word-for-word.
Given the job description text provided, if I remove things per extraction rules:
- Remove social sharing: not present.
- Remove author bylines: not present (no "by AuthorName").
- Remove "You may also like": not present.
- Remove navigation: not present.
- Remove comments: not present.
- Remove ads/promotions: not present.
- Remove cookie/subscription notices: not present.
- Remove duplicate text: maybe the title? Not sure.
- Remove category labels not part of job: maybe "World Vision Kenya:" disclaimer section? But it's part of the job listing.
- Remove stray numbers/formatting artifacts: the text has some.
If I just format as HTML per the HTML rules, I need to handle the existing tags. The text has ``, ``, ``, ``, `
`, ``, `- `. The HTML rules say "Use
for introductory paragraphs and general text", "Use
for section headers", "Use - for list items", "No ... no
tags". So I need to convert the text to use only those tags, removing
, , , ? No, "Preserve ALL original content word-for-word" means keep the text content exactly, but I can change the tags. `` and `` and `` are formatting markup, but the rule says "Preserve ALL original content word-for-word — do not add, remove, summarise, or rephrase". Does that mean keep the literal text including the tags, or keep the visible text and re-tag it? Usually in such tasks, "content" means the visible text, and you re-markup it. But it says "do not add, remove, summarise, or rephrase", which suggests don't change the words themselves.
This is tricky. Let me look at the exact wording: "Preserve ALL original content word-for-word — do not add, remove, summarise, or rephrase". If I wrap paragraphs in , I'm not removing or adding content, just tagging. If I keep , , tags, those are not "content" per se, but markup. The rule says "No CSS classes, no inline styles, no s, no
tags". It explicitly bans
tags. So I must remove
tags and presumably replace them with or just let paragraphs flow. It also says no
s. It doesn't explicitly ban , , , but says "Use for section headers", "Use - for list items". It might imply that I should use those tags for structure, and the text content goes inside them.
Given the complexity, maybe the best approach is to output the HTML version of the job description, following the HTML rules as closely as possible, while preserving all visible text word-for-word, and removing/handling the banned tags (
, , inline CSS, etc.). I'll keep , , maybe, or I can just keep the text and use etc. if needed, but the rules say "Use for section headers (e.g. Responsibilities, Qualifications, ...)". The text has `
Job Responsibilities:
` etc. I could convert those to Job Responsibilities
but the rule says "Preserve ALL original content word-for-word". If I change "Job Responsibilities:" to an , I'm not removing content, but I'm changing the tag. The content "Job Responsibilities:" would still be there.
Actually, maybe I should just wrap everything in
and
- as appropriate, and remove
by merging lines into paragraphs. The text has `
`, `
`. I'll treat each `
...
` block as a paragraph, and each `- ...
` as a list. I'll keep the text inside exactly as is. I'll remove `
` tags by not including them, and just let paragraphs be separate. I'll keep
Apply Now ↗
How well do you match?
Get an instant AI match score for this role — free, takes 3 minutes.
Tailor your CV for this role
The concierge rewrites your whole CV and writes a matching cover letter for this job — opens right here, nothing to paste.
Tailor My CV to This Job ✍️
Free cover letter for this job
Upload your CV and get a tailored cover letter in seconds — free, no account needed.
Generate a Cover Letter 📝
tags", "Preserve ALL original content word-for-word — do not add, remove, summarise, or rephrase", "If text has no clear sections, just wrap paragraphs in
tags", "Return ONLY the HTML. No markdown fences, no explanation, no preamble." I think the safest interpretation is: The user wants me to take the given job description text and format it as clean semantic HTML, following the tag rules, preserving all content word-for-word. The extraction rules might be the general persona description, but the specific task here is the HTML formatting, and the "preserve all original content word-for-word" overrides or specifies that I shouldn't remove content unless it's clearly fluff like social sharing etc. But the text doesn't have explicit social sharing, author bylines, etc. It has structural elements like `
`, ``, etc. Actually, re-reading carefully: The user message has two parts. The first part is "You are a job listing editor..." with extraction rules, and says "Return ONLY the cleaned job text." The second part starts with "You receive a plain-text job description. Your ONLY job is to format it as clean semantic HTML." and has HTML rules, and then "Job description: [text]". It's possible the user copied two different prompts, but I need to output something. Maybe the intended task is: Apply the extraction rules to clean the text, then format the cleaned text as HTML. But the HTML rules say "Preserve ALL original content word-for-word", which contradicts removing things. Unless "original content" refers to the content after extraction. Let me think differently: Perhaps the extraction rules are the primary task, and the HTML formatting is just a way to present the cleaned text. But the extraction rules say "Return ONLY the cleaned job text." not HTML. Then the second part says format as HTML. Maybe I need to do both: clean it (remove the listed items), then output the cleaned text as HTML with the specified tags, preserving the cleaned content word-for-word. Given the job description text provided, if I remove things per extraction rules: - Remove social sharing: not present. - Remove author bylines: not present (no "by AuthorName"). - Remove "You may also like": not present. - Remove navigation: not present. - Remove comments: not present. - Remove ads/promotions: not present. - Remove cookie/subscription notices: not present. - Remove duplicate text: maybe the title? Not sure. - Remove category labels not part of job: maybe "World Vision Kenya:" disclaimer section? But it's part of the job listing. - Remove stray numbers/formatting artifacts: the text has some. If I just format as HTML per the HTML rules, I need to handle the existing tags. The text has `
`, ``, ``, ``, ` for introductory paragraphs and general text", "Use , I'm not removing or adding content, just tagging. If I keep , , tags, those are not "content" per se, but markup. The rule says "No CSS classes, no inline styles, no or just let paragraphs flow. It also says no Job Responsibilities: and `. I'll treat each ` ... Get an instant AI match score for this role — free, takes 3 minutes. The concierge rewrites your whole CV and writes a matching cover letter for this job — opens right here, nothing to paste. Upload your CV and get a tailored cover letter in seconds — free, no account needed.
`, ``, `
for section headers", "Use
tags". So I need to convert the text to use only those tags, removing
, , , ? No, "Preserve ALL original content word-for-word" means keep the text content exactly, but I can change the tags. `` and `` and `` are formatting markup, but the rule says "Preserve ALL original content word-for-word — do not add, remove, summarise, or rephrase". Does that mean keep the literal text including the tags, or keep the visible text and re-tag it? Usually in such tasks, "content" means the visible text, and you re-markup it. But it says "do not add, remove, summarise, or rephrase", which suggests don't change the words themselves.
This is tricky. Let me look at the exact wording: "Preserve ALL original content word-for-word — do not add, remove, summarise, or rephrase". If I wrap paragraphs in
tags". It explicitly bans
tags. So I must remove
tags and presumably replace them with for section headers", "Use
, for section headers (e.g. Responsibilities, Qualifications, ...)". The text has `
Job Responsibilities
but the rule says "Preserve ALL original content word-for-word". If I change "Job Responsibilities:" to an , I'm not removing content, but I'm changing the tag. The content "Job Responsibilities:" would still be there.
Actually, maybe I should just wrap everything in
by merging lines into paragraphs. The text has `
`, `
` as a list. I'll keep the text inside exactly as is. I'll remove `
` tags by not including them, and just let paragraphs be separate. I'll keepHow well do you match?
Tailor your CV for this role
Free cover letter for this job
Let jobs find you
Leave your email and our AI matches you to new jobs across 24 African markets — free. You wait for the call.
Add your CV for real matches
Upload your CV and we score every new job against your real experience — only strong matches reach your inbox. Optional, but it makes your matches far sharper.
You're in.
We'll email you the moment a job matches your profile. Check your inbox for a welcome from My Job Concierge.
I'm ECHO, your MJC career assistant. I can help you find jobs, explore career tools, and connect with opportunities across Africa.