[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"me":3,"catalog:en:data-engineer":4,"config":126},null,{"field_key":5,"field_name":6,"seniority":7,"topic_key":7,"topic_name":7,"spec_key":7,"spec_name":7,"locale":8,"cell_total":9,"field_total":9,"seniorities":10,"topics":14,"specs":40,"samples":41},"data-engineer","Data Engineer","","en",600,[11,12,13],"junior","mid","senior",[15,19,22,25,28,31,34,37],{"key":16,"name":17,"count":18},"data-governance-lineage","Data Governance Lineage",75,{"key":20,"name":21,"count":18},"data-modeling-warehousing","Data Modeling Warehousing",{"key":23,"name":24,"count":18},"data-partitioning-scaling","Data Partitioning Scaling",{"key":26,"name":27,"count":18},"data-pipeline-design","Data Pipeline Design",{"key":29,"name":30,"count":18},"data-pipeline-orchestration","Data Pipeline Orchestration",{"key":32,"name":33,"count":18},"data-pipeline-reliability","Data Pipeline Reliability",{"key":35,"name":36,"count":18},"data-quality-validation","Data Quality Validation",{"key":38,"name":39,"count":18},"streaming-fundamentals","Streaming Fundamentals",[],[42,60,74,87,100,113],{"id":43,"topic":17,"difficulty":44,"body":45,"options":46,"correct_key":51,"explanation":59},"019f6a40-d699-7521-965a-6956978e9b32",1,"In the context of a data platform, what best describes 'data lineage'?",[47,50,53,56],{"key":48,"text":49},"a","The physical storage location where a dataset's files are kept on disk",{"key":51,"text":52},"b","A traceable record of data's origin and transformations downstream",{"key":54,"text":55},"c","A naming convention applied to tables and columns",{"key":57,"text":58},"d","The access permissions assigned to a dataset's owner","Data lineage is the record of where data originates and how it's transformed on its way to consumption; it is not a backup schedule, naming convention, or permission list.",{"id":61,"topic":17,"difficulty":62,"body":63,"options":64,"correct_key":57,"explanation":73},"019f6a40-d6b2-76bc-bfb5-53c7f02c931a",2,"A dashboard metric suddenly shows an unexpected drop. Before touching any code, what is the most useful first step that lineage information enables?",[65,67,69,71],{"key":48,"text":66},"Increasing the retry count of the job that produces the dashboard",{"key":51,"text":68},"Asking the dashboard's end users to describe what they expected to see",{"key":54,"text":70},"Restarting the entire data platform to clear any cached state",{"key":57,"text":72},"Tracing the metric backward through lineage to the changed source","Lineage lets an engineer trace the metric backward to the exact upstream source or transformation that changed, instead of guessing via retries or restarts.",{"id":75,"topic":17,"difficulty":44,"body":76,"options":77,"correct_key":48,"explanation":86},"019f6a40-d6b4-71a2-9adc-d4dcc82e1a71","Which statement about data lineage is accurate?",[78,80,82,84],{"key":48,"text":79},"Lineage can be captured at different levels, e.g. table or column",{"key":51,"text":81},"Lineage only matters for data stored in spreadsheets",{"key":54,"text":83},"Lineage is generated once and never needs to be updated",{"key":57,"text":85},"Lineage information is identical to a database's index structure","Lineage can be captured at different granularities, from whole tables down to individual columns, depending on the need.",{"id":88,"topic":17,"difficulty":62,"body":89,"options":90,"correct_key":54,"explanation":99},"019f6a40-d6b4-7eaf-9a02-83ed1dd27340","A team plans to drop a column from a shared table. What should they check first to avoid breaking anything?",[91,93,95,97],{"key":48,"text":92},"Whether the column name is spelled consistently across the codebase",{"key":51,"text":94},"How much disk space the column currently occupies",{"key":54,"text":96},"The lineage graph, to see which downstream consumers use it",{"key":57,"text":98},"The color scheme used in dashboards that display the table","The lineage graph shows exactly which downstream reports and pipelines depend on that column, so the team can assess the real impact before dropping it.",{"id":101,"topic":17,"difficulty":62,"body":102,"options":103,"correct_key":51,"explanation":112},"019f6a40-d6b5-79a7-bbcf-f465a118826d","Why is column-level lineage generally more useful than table-level lineage for impact analysis?",[104,106,108,110],{"key":48,"text":105},"It requires less metadata to maintain overall",{"key":51,"text":107},"It shows exactly which downstream fields depend on that column",{"key":54,"text":109},"It automatically fixes broken pipelines without manual intervention",{"key":57,"text":111},"It removes the need for a data catalog entirely","Column-level lineage pinpoints which specific downstream fields are affected, avoiding wasted effort investigating unrelated parts of a table.",{"id":114,"topic":17,"difficulty":62,"body":115,"options":116,"correct_key":57,"explanation":125},"019f6a40-d6b6-735a-ad88-477bb10a4ca3","Before renaming a widely-used table, an engineer wants to know exactly which reports would break. Lineage tracking helps by...",[117,119,121,123],{"key":48,"text":118},"Automatically renaming the table in every downstream system",{"key":51,"text":120},"Notifying all employees in the company by broadcast email",{"key":54,"text":122},"Encrypting the table so unauthorized changes are prevented",{"key":57,"text":124},"Listing every downstream consumer of that table","Lineage tracking lists every downstream consumer of a table, so each one can be reviewed or updated before a breaking change like a rename.",{"fields":127,"seniorities":311,"interview_shapes":312,"locales":317,"oauth":319,"question_count":322,"coach_enabled":323,"jd_match_enabled":323},[128,153,173,190,214,227,246,265,287,294,298,305],{"key":129,"name_tr":130,"name_en":130,"sort":44,"specializations":131},"backend","Backend",[132,135,138,141,144,147,150],{"key":133,"name":134,"field":129},"general","Genel",{"key":136,"name":137,"field":129},"go","Go",{"key":139,"name":140,"field":129},"python","Python",{"key":142,"name":143,"field":129},"java","Java",{"key":145,"name":146,"field":129},"csharp","C#\u002F.NET",{"key":148,"name":149,"field":129},"nodejs","Node.js",{"key":151,"name":152,"field":129},"php","PHP",{"key":154,"name_tr":155,"name_en":155,"sort":62,"specializations":156},"frontend","Frontend",[157,158,161,164,167,170],{"key":133,"name":134,"field":154},{"key":159,"name":160,"field":154},"javascript","JavaScript",{"key":162,"name":163,"field":154},"typescript","TypeScript",{"key":165,"name":166,"field":154},"react","React",{"key":168,"name":169,"field":154},"vue","Vue",{"key":171,"name":172,"field":154},"angular","Angular",{"key":174,"name_tr":175,"name_en":175,"sort":176,"specializations":177},"fullstack","Fullstack",3,[178,179,180,181,182,183,184,185,186,187,188,189],{"key":133,"name":134,"field":174},{"key":136,"name":137,"field":129},{"key":139,"name":140,"field":129},{"key":142,"name":143,"field":129},{"key":145,"name":146,"field":129},{"key":148,"name":149,"field":129},{"key":151,"name":152,"field":129},{"key":159,"name":160,"field":154},{"key":162,"name":163,"field":154},{"key":165,"name":166,"field":154},{"key":168,"name":169,"field":154},{"key":171,"name":172,"field":154},{"key":191,"name_tr":192,"name_en":192,"sort":193,"specializations":194},"devops-cloud","DevOps \u002F Cloud",4,[195,196,199,202,205,208,211],{"key":133,"name":134,"field":191},{"key":197,"name":198,"field":191},"aws","AWS",{"key":200,"name":201,"field":191},"gcp","GCP",{"key":203,"name":204,"field":191},"azure","Azure",{"key":206,"name":207,"field":191},"kubernetes","Kubernetes",{"key":209,"name":210,"field":191},"terraform","Terraform",{"key":212,"name":213,"field":191},"linux","Linux",{"key":215,"name_tr":216,"name_en":216,"sort":217,"specializations":218},"ai-engineer","AI Engineer",5,[219,220,221,224],{"key":133,"name":134,"field":215},{"key":139,"name":140,"field":215},{"key":222,"name":223,"field":215},"llm-rag","LLM\u002FRAG",{"key":225,"name":226,"field":215},"mlops","MLOps",{"key":228,"name_tr":229,"name_en":230,"sort":231,"specializations":232},"database","Veritabanı","Database",6,[233,234,237,240,243],{"key":133,"name":134,"field":228},{"key":235,"name":236,"field":228},"postgresql","PostgreSQL",{"key":238,"name":239,"field":228},"mysql","MySQL",{"key":241,"name":242,"field":228},"mongodb","MongoDB",{"key":244,"name":245,"field":228},"redis","Redis",{"key":247,"name_tr":248,"name_en":249,"sort":250,"specializations":251},"mobile","Mobil","Mobile",7,[252,253,256,259,262],{"key":133,"name":134,"field":247},{"key":254,"name":255,"field":247},"ios-swift","iOS (Swift)",{"key":257,"name":258,"field":247},"android-kotlin","Android (Kotlin)",{"key":260,"name":261,"field":247},"flutter","Flutter",{"key":263,"name":264,"field":247},"react-native","React Native",{"key":266,"name_tr":267,"name_en":268,"sort":269,"specializations":270},"security","Güvenlik","Security",8,[271,272,275,278,281,284],{"key":133,"name":134,"field":266},{"key":273,"name":274,"field":266},"appsec","AppSec",{"key":276,"name":277,"field":266},"offensive-pentest","Offensive \u002F Pentest",{"key":279,"name":280,"field":266},"cloud-security","Cloud Security",{"key":282,"name":283,"field":266},"devsecops","DevSecOps",{"key":285,"name":286,"field":266},"blue-team-incident","Blue Team \u002F Incident",{"key":288,"name_tr":289,"name_en":290,"sort":291,"specializations":292},"qa-test-automation","QA \u002F Test Otomasyonu","QA \u002F Test Automation",9,[293],{"key":133,"name":134,"field":288},{"key":5,"name_tr":6,"name_en":6,"sort":295,"specializations":296},10,[297],{"key":133,"name":134,"field":5},{"key":299,"name_tr":300,"name_en":301,"sort":302,"specializations":303},"game-dev","Oyun Geliştirme","Game Development",11,[304],{"key":133,"name":134,"field":299},{"key":306,"name_tr":307,"name_en":307,"sort":308,"specializations":309},"ml-engineer","ML Engineer",12,[310],{"key":133,"name":134,"field":306},[11,12,13],{"junior":313,"mid":315,"senior":316},{"questions":314,"median_sec":3},20,{"questions":314,"median_sec":3},{"questions":314,"median_sec":3},[318,8],"tr",[320,321],"google","github",21750,true]